2 дня назад
Site Reliability Engineer (Azure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Azure): Maintaining and improving the availability, performance, and reliability of critical multi-cloud SaaS applications with an accent on Azure infrastructure, automation, monitoring, and incident response. Focus on troubleshooting AKS and networking failures, designing disaster recovery and multi-region failover, and driving root-cause remediation across the full incident lifecycle.
Location: Home Office, Philippines
Company
provides a cloud-native Identity Security Platform for securing human and machine identities across modern enterprise infrastructure.
What you will do
- Own the availability and performance of production SaaS applications running on Azure across multiple geographic regions.
- Troubleshoot and resolve AKS, deployment, ingress, networking, autoscaling, and cloud infrastructure issues.
- Participate in an on-call rotation, including weekends, and lead incident response from detection through resolution.
- Improve disaster recovery, failover, observability, reliability, and performance practices for multi-region deployments.
- Build automation scripts and monitoring tools using operational data to reduce manual work.
- Write post-incident reviews, identify root causes, drive preventive actions, and communicate incident updates to customer-facing stakeholders.
Requirements
- 5+ years of experience in Site Reliability Engineering, DevOps, or cloud administration with ownership of production systems.
- Hands-on Azure administration, including AKS, core Azure services, cloud networking, and cloud security fundamentals.
- Experience with monitoring, logging, alerting, Datadog APM, Azure Monitor, or ELK, including log and stack-trace analysis.
- Knowledge of firewalls, load balancers, VPNs, DNS, routing, backup, redundancy, and cloud disaster recovery.
- Automation and scripting experience with PowerShell, Python, or similar tools.
- Clear written communication, customer-first incident management, and ownership from detection through post-mortem.
Nice to have
- Experience with AWS Cloud Platform.
- Experience with Azure DevOps and infrastructure-as-code tools such as Terraform or ARM templates.
- Experience operating SaaS products with regional tenant architectures and multiple geo-specific production environments.
Culture & Benefits
- Meaningful work securing human and machine identities for large organizations.
- Collaborative, innovative environment centered on respect, trust, ownership, and adaptability.
- Career progression and support for self-motivated, collaborative employees.
- Healthcare insurance, pension or retirement matching, life insurance, employee assistance, paid time off, and company holidays.
- Competitive salary and a bonus program.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 часов назад
Site Reliability Engineer (Azure)
11 часов назад
Lead Site Reliability Engineer (Azure)
6 дней назад
Site Reliability Engineer - DevSecOps Engineer (Cloud)
7 дней назад
Senior DevOps Engineer (Azure)
10 часов назад
Senior DevOps Engineer (Azure)
8 часов назад
Senior DevOps Engineer / Site Reliability Engineer (AI)
170 000 - 220 000$