Назад
Company hidden
3 часа назад

Senior Manager, Infrastructure Platform Engineering (AI)

245 000 - 295 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Manager, Infrastructure Platform Engineering (AI): Building reliable, secure, and efficiently allocatable platform systems for large-scale compute infrastructure with an accent on resource allocation, capacity intelligence, lifecycle management, and platform security. Focus on leading infrastructure software engineers, designing Kubernetes-based distributed systems, reducing stranded resources, and improving platform reliability across cloud and on-premise environments.

Location: On-site in San Francisco or Sunnyvale, California, US

Salary: $245,000–$295,000 annually, plus bonus and restricted stock units

Company

hirify.global is an AI infrastructure company operating an integrated stack spanning energy, data center infrastructure, and cloud services for large-scale AI workloads.

What you will do

  • Lead the team building platform services for resource pooling, allocation, fleet state tracking, and infrastructure lifecycle management.
  • Set the technical roadmap for capacity and utilization intelligence, resource reconciliation, and platform security and trust.
  • Drive the design of secure, observable platform systems using Kubernetes orchestration, automation, and lower-level system and hardware integrations.
  • Hire, mentor, and develop infrastructure software engineers while establishing a high-performing engineering organization.
  • Partner with infrastructure, production engineering, and security teams to improve reliability, capacity, availability, and operational recovery.
  • Remain technically hands-on through design reviews, architecture decisions, engineering standards, and deployment and on-call practices.

Requirements

  • 10+ years of experience in infrastructure or systems software development, including 3+ years in engineering leadership.
  • Deep expertise building large-scale infrastructure platforms that pool, allocate, and reconcile compute resources.
  • Strong experience with Kubernetes, cloud platforms such as GCP, AWS, or Azure, orchestration, automation, and distributed systems in production.
  • Experience with distributed state management, resource and system lifecycle modeling, failure handling, and desired-versus-actual state reconciliation.
  • Experience in efficiency, capacity, or performance engineering, including bottleneck analysis and measurable improvements in utilization or availability.
  • Player-coach management experience, including hiring and developing infrastructure engineers in ambiguous, fast-moving environments.

Nice to have

  • Experience operating Kubernetes on bare-metal infrastructure and managed services such as GKE, EKS, or AKS.
  • Experience with GPU clusters, AI training and inference workloads, capacity forecasting, demand modeling, or allocation optimization.
  • Knowledge of secure boot, measured boot, TPMs, hardware attestation, or hardware-software co-design.
  • Experience with Prometheus, OpenTelemetry, Grafana, hyperscale cloud platforms, or infrastructure products serving internal engineering teams.

Culture & Benefits

  • Paid time off, paid holidays, leave programs, and paid parental leave.
  • Health, dental, and vision insurance, HSA contributions, life insurance, disability coverage, and mental health support.
  • Professional development, tuition reimbursement, volunteer time off, and commuter benefits.
  • 401(k) retirement plan with company matching up to 4% of salary.
  • Daily meal allowance, cell phone stipend, global travel insurance, and location-specific programs.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →