Назад
Company hidden
6 дней назад

Staff DevOps Engineer (Kubernetes)

Формат работы
remote (только Mexico)/hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US/Mexico
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff DevOps Engineer (Kubernetes): Leading the architecture and evolution of a multi-cluster Kubernetes platform and internal developer platform with an accent on GitOps, cloud infrastructure, reliability, and AI/ML workloads. Focus on designing service mesh and networking strategies, implementing GPU scheduling for training and inference, optimizing capacity and cost, and defining platform standards across engineering.

Location: Remote across Mexico. Team members within 80 kilometers of the Guadalajara office are expected to work onsite.

Company

hirify.global provides an AI-powered, conversation-centric customer experience and team collaboration platform.

What you will do

  • Own the architecture and roadmap for a multi-cluster Kubernetes platform, including scaling, upgrades, and multi-tenancy.
  • Establish GitOps deployment workflows and build an internal developer platform with self-service access to compute, environments, and observability.
  • Design service mesh, networking, ingress, and secure communication strategies for a redundant microservices architecture.
  • Define GPU scheduling, node pools, quotas, and resource management for AI/ML training and inference workloads.
  • Lead capacity planning, reliability engineering, SLOs, incident response, postmortems, and cloud cost optimization.
  • Set technical direction, review platform-impacting architecture, mentor engineers, and influence standards across teams.

Requirements

  • 8+ years of DevOps, platform, or infrastructure engineering experience, including 5+ years operating Kubernetes in production.
  • Hands-on experience with AWS and Google Cloud Platform, Terraform or Pulumi, and GitOps using Argo CD or Flux.
  • Deep knowledge of container orchestration, networking, ingress, service mesh, Linux, storage, and security fundamentals.
  • Experience with Nginx, Kafka, Redis, and observability platforms such as Prometheus, Grafana, Datadog, or OpenTelemetry.
  • Experience supporting GPU and ML/AI workloads in GKE and EKS, plus internal developer platforms or paved-road tooling.
  • Bachelor’s degree in Computer Science or a related field, or equivalent experience; strong communication and cross-team influence skills.

Nice to have

  • Kubernetes policy-as-code and security tooling such as OPA/Gatekeeper, Kyverno, or image scanning.
  • AWS, GCP, or Azure certifications, or CKA/CKS certification.

Culture & Benefits

  • Health, vision, and dental insurance for employees and eligible family members.
  • Life insurance, personal days, vacation days, vacation premium, food vouchers, and a matched savings fund.
  • Employee Assistance Program and wellness initiatives.
  • Ongoing learning, development opportunities, and career advancement.
  • Results-driven, collaborative, and customer-focused working environment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →