обновлено 3 часа назад
Site Reliability Engineer (Azure/AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Azure/AI): Building end-to-end observability and reliability solutions for mission-critical payment and Order to Cash workflows with an accent on Azure Monitor, ADX/KQL, Grafana, and AI-assisted diagnostics. Focus on defining SLIs and SLOs, automating incident prevention and self-healing, analyzing telemetry quality, and optimizing capacity and performance.
Location: Buenos Aires, Buenos Aires, Argentina
Company
Global Business Services operates a multifunction service center in Buenos Aires that delivers business services and technology solutions globally.
What you will do
- Define, implement, and monitor SLIs, SLOs, and error budgets for mission-critical payment and Order to Cash workflows.
- Architect, deploy, and maintain scalable observability solutions across distributed environments using Azure Monitor, ADX/KQL, and Grafana.
- Analyze incidents, facilitate blameless root cause analysis, and eliminate observability blind spots.
- Automate repetitive operational tasks and develop self-healing mechanisms.
- Design AI-assisted diagnostics, smart alerting, and anomaly detection solutions.
- Support load testing, capacity planning, system health monitoring, and performance optimization.
Requirements
- Strong experience in production reliability, incident response, and system architecture.
- Hands-on proficiency with Python, PowerShell, and C# or another backend language.
- Advanced experience with Grafana dashboards, ADX/KQL log analysis, and telemetry validation.
- Experience collaborating with engineering teams on code-level issues and translating technical metrics into business impact.
- Practical experience using AI tools such as M365 or GitHub Copilot and building AI agents.
- Visa sponsorship is not available, and applicants with time-limited visa status will not be considered.
Nice to have
- Experience with Infrastructure as Code and CI/CD pipelines.
Culture & Benefits
- Work within a vendor-driven operations team focused on reliability and incident prevention.
- Collaborate across engineering and business functions.
- Relocation may be considered.
- Expatriate assignments are not considered.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →