Назад
Company hidden
2 дня назад

Site Reliability Engineer (Azure)

Формат работы
remote (только Philippines)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Philippines
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (Azure): Maintaining and improving the availability, performance, and reliability of critical multi-cloud SaaS applications with an accent on Azure infrastructure, automation, monitoring, and incident response. Focus on troubleshooting AKS and networking failures, designing disaster recovery and multi-region failover, and driving root-cause remediation across the full incident lifecycle.

Location: Home Office, Philippines

Company

hirify.global provides a cloud-native Identity Security Platform for securing human and machine identities across modern enterprise infrastructure.

What you will do

  • Own the availability and performance of production SaaS applications running on Azure across multiple geographic regions.
  • Troubleshoot and resolve AKS, deployment, ingress, networking, autoscaling, and cloud infrastructure issues.
  • Participate in an on-call rotation, including weekends, and lead incident response from detection through resolution.
  • Improve disaster recovery, failover, observability, reliability, and performance practices for multi-region deployments.
  • Build automation scripts and monitoring tools using operational data to reduce manual work.
  • Write post-incident reviews, identify root causes, drive preventive actions, and communicate incident updates to customer-facing stakeholders.

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps, or cloud administration with ownership of production systems.
  • Hands-on Azure administration, including AKS, core Azure services, cloud networking, and cloud security fundamentals.
  • Experience with monitoring, logging, alerting, Datadog APM, Azure Monitor, or ELK, including log and stack-trace analysis.
  • Knowledge of firewalls, load balancers, VPNs, DNS, routing, backup, redundancy, and cloud disaster recovery.
  • Automation and scripting experience with PowerShell, Python, or similar tools.
  • Clear written communication, customer-first incident management, and ownership from detection through post-mortem.

Nice to have

  • Experience with AWS Cloud Platform.
  • Experience with Azure DevOps and infrastructure-as-code tools such as Terraform or ARM templates.
  • Experience operating SaaS products with regional tenant architectures and multiple geo-specific production environments.

Culture & Benefits

  • Meaningful work securing human and machine identities for large organizations.
  • Collaborative, innovative environment centered on respect, trust, ownership, and adaptability.
  • Career progression and support for self-motivated, collaborative employees.
  • Healthcare insurance, pension or retirement matching, life insurance, employee assistance, paid time off, and company holidays.
  • Competitive salary and a bonus program.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →