6 дней назад
Staff DevOps Engineer (Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff DevOps Engineer (Kubernetes): Leading the architecture and evolution of a multi-cluster Kubernetes platform and internal developer platform with an accent on GitOps, cloud infrastructure, reliability, and AI/ML workloads. Focus on designing service mesh and networking strategies, implementing GPU scheduling for training and inference, optimizing capacity and cost, and defining platform standards across engineering.
Location: Remote across Mexico. Team members within 80 kilometers of the Guadalajara office are expected to work onsite.
Company
provides an AI-powered, conversation-centric customer experience and team collaboration platform.
What you will do
- Own the architecture and roadmap for a multi-cluster Kubernetes platform, including scaling, upgrades, and multi-tenancy.
- Establish GitOps deployment workflows and build an internal developer platform with self-service access to compute, environments, and observability.
- Design service mesh, networking, ingress, and secure communication strategies for a redundant microservices architecture.
- Define GPU scheduling, node pools, quotas, and resource management for AI/ML training and inference workloads.
- Lead capacity planning, reliability engineering, SLOs, incident response, postmortems, and cloud cost optimization.
- Set technical direction, review platform-impacting architecture, mentor engineers, and influence standards across teams.
Requirements
- 8+ years of DevOps, platform, or infrastructure engineering experience, including 5+ years operating Kubernetes in production.
- Hands-on experience with AWS and Google Cloud Platform, Terraform or Pulumi, and GitOps using Argo CD or Flux.
- Deep knowledge of container orchestration, networking, ingress, service mesh, Linux, storage, and security fundamentals.
- Experience with Nginx, Kafka, Redis, and observability platforms such as Prometheus, Grafana, Datadog, or OpenTelemetry.
- Experience supporting GPU and ML/AI workloads in GKE and EKS, plus internal developer platforms or paved-road tooling.
- Bachelor’s degree in Computer Science or a related field, or equivalent experience; strong communication and cross-team influence skills.
Nice to have
- Kubernetes policy-as-code and security tooling such as OPA/Gatekeeper, Kyverno, or image scanning.
- AWS, GCP, or Azure certifications, or CKA/CKS certification.
Culture & Benefits
- Health, vision, and dental insurance for employees and eligible family members.
- Life insurance, personal days, vacation days, vacation premium, food vouchers, and a matched savings fund.
- Employee Assistance Program and wellness initiatives.
- Ongoing learning, development opportunities, and career advancement.
- Results-driven, collaborative, and customer-focused working environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Software Engineer, Infrastructure & Platform (Aerospace)
8 дней назад
DevOps Engineer (Terraform)
120 000 - 130 000$
5 дней назад
Senior Software Engineer (Infrastructure)
170 000 - 220 000$
8 дней назад
Sr. Staff Infrastructure Engineer (AWS)
7 дней назад
DevOps Engineer (AI)
163 000 - 204 000$
8 дней назад