Назад
1 день назад

Platform Engineer (AI)

230 000 - 340 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Platform Engineer (Kubernetes/Go): Architecting and operating Kubernetes clusters across AWS and bare-metal datacenters with an accent on automation, reliability, and scalability. Focus on building cluster lifecycle management, implementing observability, and designing cloud-native services for AI infrastructure.

Location: Hybrid (San Francisco, San Jose, or Seattle/Bellevue). Presence required in office 4 days per week

Salary: $230,000 – $340,000

Company

Lambda is a leader in AI cloud infrastructure, providing supercomputing power to AI researchers and enterprises.

What you will do

  • Architect, deploy, and operate Kubernetes clusters across AWS and bare-metal datacenters.
  • Build and maintain automation for cluster provisioning, upgrades, and scaling.
  • Own the reliability, performance, and security of production Kubernetes workloads.
  • Implement observability, logging, and alerting for clusters and critical workloads.
  • Collaborate with product teams to design scalable cloud-native services and CI/CD pipelines.
  • Lead incident response, root-cause analysis, and mentor other engineers.

Requirements

  • 5+ years in Platform, Infrastructure, or SRE roles.
  • Experience running Kubernetes in production at scale, including deep knowledge of internals.
  • Proficiency with Helm, Kustomize, and GitOps-based delivery.
  • Strong skills in Infrastructure-as-Code (Terraform, Pulumi, or equivalent).
  • Hands-on experience with observability stacks (Prometheus, Grafana, OpenTelemetry).
  • Strong coding skills in Go or Python for automation and tooling.

Nice to have

  • Experience with multi-cluster, multi-cloud, or hybrid environments.
  • Knowledge of GPU scheduling, HPC workloads, or ML/AI infrastructure.
  • Experience with workflow orchestration (Temporal, Cadence, or Argo Workflows).
  • CKA/CKS certifications.

Culture & Benefits

  • Generous cash and equity compensation.
  • Comprehensive health, dental, and vision coverage for employees and dependents.
  • 401k plan with 2% company match.
  • Wellness and commuter stipends for select roles.
  • Flexible paid time off plan.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →