Назад
3 часа назад

Engineering Manager (Kubernetes Infrastructure)

182 000 - 242 000$
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Engineering Manager (Kubernetes Infrastructure): Leading the team building and operating bare-metal Kubernetes infrastructure for high-performance AI and ML workloads with an accent on cluster lifecycle management, reliability, observability, and infrastructure automation. Focus on improving incident response, rollout safety, provisioning, upgrade reliability, and cross-functional execution across compute, networking, storage, security, and product teams.

Location: Livingston, NJ; New York, NY; Sunnyvale, CA; or Bellevue, WA

Salary: $182,000–$242,000 base salary annually, plus discretionary bonus, equity awards, and benefits.

Company

CoreWeave provides cloud infrastructure, tools, and services for building and scaling AI workloads on high-performance computing infrastructure.

What you will do

  • Lead a team responsible for Kubernetes infrastructure running directly on bare metal.
  • Set team goals, priorities, execution plans, and operating mechanisms.
  • Partner with senior engineers and platform teams on cluster lifecycle management, upgrades, reliability, observability, and automation.
  • Improve incident response, on-call health, root-cause analysis, service ownership, and production readiness.
  • Support provisioning, patching, scaling, and troubleshooting of Kubernetes clusters.
  • Hire, coach, and develop engineers while building strong relationships with compute, networking, storage, security, and product teams.

Requirements

  • Experience managing an infrastructure, platform, or SRE-oriented engineering team.
  • Strong technical depth in Kubernetes, distributed systems, and production infrastructure.
  • Experience operating Kubernetes in complex environments, including bare metal, hybrid, or performance-sensitive systems.
  • Experience with cluster provisioning, upgrades, node operations, observability, reliability engineering, and incident response.
  • Ability to improve execution, engineering quality, operational maturity, and cross-functional collaboration.
  • Applicants must meet applicable U.S. export-control access requirements, including U.S. person status or eligibility for required authorization.

Nice to have

  • Experience with GPU-heavy, HPC, or ML infrastructure.
  • Experience with bare-metal infrastructure, server lifecycle operations, or low-level systems troubleshooting.
  • Familiarity with Kubernetes networking, storage, and security primitives.
  • Experience building internal platform products and automating infrastructure with Go, Python, controllers, operators, or configuration management systems.

Culture & Benefits

  • Medical, dental, and vision insurance fully paid by CoreWeave, plus life and disability insurance.
  • 401(k) with employer match, flexible spending and health savings accounts, tuition reimbursement, and an employee stock purchase program.
  • Paid parental leave, family-forming support, childcare support, flexible PTO, and mental wellness benefits.
  • Collaborative, entrepreneurial work environment focused on innovation, independent thinking, and high-performance client experiences.
  • Catered lunches at office and data center locations.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →