2 дня назад
Senior SRE Engineer (AWS/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior SRE Engineer (AWS/Kubernetes): Building and operating scalable cloud infrastructure, deployment automation, and observability systems for Flowcode's platforms with an accent on reliability, resilience, and high-availability production environments. Focus on managing EKS-based infrastructure, designing GitHub Actions and ArgoCD delivery pipelines, engineering SLOs, and improving disaster recovery and incident response.
Location: New York, NY (Hybrid) / Remote
Company
develops and operates digital platforms supported by scalable cloud infrastructure and reliability engineering.
What you will do
- Improve availability, scalability, and resilience across 's platforms.
- Own and operate EKS-based infrastructure and core AWS services, including VPC and RDS.
- Design and scale CI/CD pipelines with GitHub Actions, GitOps workflows with ArgoCD, and Helm-based deployments.
- Strengthen disaster recovery, failover mechanisms, and progressive delivery with automated validation.
- Build monitoring, logging, alerting, metrics, tracing, and visualization dashboards while establishing SLOs.
- Participate in incident response and postmortems, converting findings into durable fixes and supporting engineering teams with infrastructure issues.
Requirements
- 4+ years of professional experience in SRE, DevOps, or Platform Engineering.
- Strong Kubernetes expertise, including cluster troubleshooting and management of controllers or CRDs.
- Advanced Terraform or OpenTofu skills, including module architecture and production state management.
- Hands-on experience with ArgoCD, Helm, GitHub Actions, and production CI/CD automation.
- Ability to write production-grade Go or Python and robust shell scripts.
- Expertise in AWS, distributed systems, high-availability production environments, infrastructure initiatives, and operational support.
Nice to have
- Experience with Crossplane or other Kubernetes-native infrastructure provisioning solutions.
- Deep observability experience with Datadog or Prometheus, including SLOs, dashboards, and alerting.
- Knowledge of secrets management frameworks.
- Experience with Karpenter, Cluster Autoscaler, or other cluster-efficiency technologies.
Culture & Benefits
- Individual-contributor role within the engineering organization.
- Work in a high-performance environment with significant infrastructure ownership and impact.
- Equal opportunity workplace committed to diversity and non-discrimination.
- Starting pay is determined by role, skills, experience, qualifications, work location, and market conditions.
Hiring process
- Apply for consideration; compensation and final offer details are determined based on role-related qualifications and work location.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
17 часов назад
Senior Software Engineer (Kubernetes)
162 000 - 260 000$
3 дня назад
Senior DevOps Engineer, AI Platform
4 дня назад
Senior DevOps Engineer (AWS/Kubernetes)
140 000 - 165 000$
5 дней назад
Sr. Site Reliability Engineer (AWS/CI/CD)
3 дня назад
Senior DevOps Engineer
4 дня назад
Senior AWS DevOps Engineer
91 200 - 120 000$