2 дня назад
Senior Site Reliability Engineer (AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AWS) (cloud infrastructure and platform reliability): Designing, building, and operating highly reliable, scalable, and secure platforms across hybrid on-premises and AWS environments with an accent on automation, resiliency, observability, and cost efficiency. Focus on multi-region disaster recovery, infrastructure as code, self-healing systems, performance optimization, and mission-critical service operations.
Location: Hybrid, Manila — 6805 Ayala Ave
Company
provides business-critical technology platforms and fosters a collaborative, inclusive workplace.
What you will do
- Design and operate highly available, fault-tolerant platforms across on-premises and AWS environments.
- Lead multi-region disaster recovery planning, testing, and validation of RTO/RPO targets.
- Build Terraform-based infrastructure, reusable modules, deployment pipelines, guardrails, and self-healing automation.
- Implement end-to-end observability with metrics, logs, traces, synthetic monitoring, dashboards, alerts, and runbooks.
- Engineer AWS platforms using services including EKS, EC2, RDS/Aurora, Lambda, API Gateway, CloudFront, WAF, ALB/NLB, CloudWatch, X-Ray, IAM, and Secrets Manager.
- Improve performance, scalability, deployment safety, security, compliance, and infrastructure cost efficiency.
Requirements
- 8+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Systems Engineering.
- Hands-on expertise with AWS, cloud-native architectures, Terraform, and Infrastructure as Code.
- Strong programming experience in Python, Java, or similar languages.
- Deep experience with Linux/Unix systems and strong knowledge of networking, security, and distributed systems.
- Experience operating mission-critical, high-volume platforms.
- Availability for the hybrid role in Manila.
Nice to have
- Experience in financial services or other highly regulated environments.
- Experience with EKS/Kubernetes at scale.
- Familiarity with Chaos Engineering and resilience testing.
- Experience leading FinOps and cloud cost-optimization initiatives.
- Experience transitioning traditional infrastructure teams to SRE practices.
Culture & Benefits
- Collaborative and engaging work environment.
- Inclusive workplace that values authenticity and diverse perspectives.
- Human review is included in employment decisions when AI-based recruiting tools are used.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →