Назад
Company hidden
обновлено 3 часа назад

Site Reliability Engineer (Azure/AI)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Argentina
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (Azure/AI): Building end-to-end observability and reliability solutions for mission-critical payment and Order to Cash workflows with an accent on Azure Monitor, ADX/KQL, Grafana, and AI-assisted diagnostics. Focus on defining SLIs and SLOs, automating incident prevention and self-healing, analyzing telemetry quality, and optimizing capacity and performance.

Location: Buenos Aires, Buenos Aires, Argentina

Company

hirify.global Global Business Services operates a multifunction service center in Buenos Aires that delivers business services and technology solutions globally.

What you will do

  • Define, implement, and monitor SLIs, SLOs, and error budgets for mission-critical payment and Order to Cash workflows.
  • Architect, deploy, and maintain scalable observability solutions across distributed environments using Azure Monitor, ADX/KQL, and Grafana.
  • Analyze incidents, facilitate blameless root cause analysis, and eliminate observability blind spots.
  • Automate repetitive operational tasks and develop self-healing mechanisms.
  • Design AI-assisted diagnostics, smart alerting, and anomaly detection solutions.
  • Support load testing, capacity planning, system health monitoring, and performance optimization.

Requirements

  • Strong experience in production reliability, incident response, and system architecture.
  • Hands-on proficiency with Python, PowerShell, and C# or another backend language.
  • Advanced experience with Grafana dashboards, ADX/KQL log analysis, and telemetry validation.
  • Experience collaborating with engineering teams on code-level issues and translating technical metrics into business impact.
  • Practical experience using AI tools such as M365 or GitHub Copilot and building AI agents.
  • Visa sponsorship is not available, and applicants with time-limited visa status will not be considered.

Nice to have

  • Experience with Infrastructure as Code and CI/CD pipelines.

Culture & Benefits

  • Work within a vendor-driven operations team focused on reliability and incident prevention.
  • Collaborate across engineering and business functions.
  • Relocation may be considered.
  • Expatriate assignments are not considered.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →