Назад
Company hidden
2 дня назад

Technical Program Manager (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Technical Program Manager (AI inference and capacity planning): Leading capacity planning, fleet strategy, and infrastructure readiness for a large-scale AI inference service with an accent on forecasting, cluster allocation, and utilization. Focus on coordinating model deployments, improving capacity management tools, and resolving bottlenecks across Engineering, Product, Infrastructure, SRE, Operations, and executive leadership.

Location: Hybrid, with offices in Sunnyvale and Toronto, plus a remote office

Company

hirify.global builds large-scale AI hardware and inference infrastructure designed to deliver high-speed model training and inference.

What you will do

  • Build and maintain 6-, 12-, and 26-week capacity forecasts across inference clusters, translating customer contracts and sales pipeline needs into infrastructure requirements.
  • Coordinate new datacenter capacity bring-up, production readiness, engineering delivery, and related automation.
  • Lead weekly capacity reviews, model placement, cluster rebalancing, deployment planning, and fleet utilization reporting.
  • Drive adoption and continuous improvement of the internal capacity planning and allocation platform, including testing, issue resolution, pilots, and deployments.
  • Identify capacity bottlenecks, risks, and dependencies; coordinate incident resolution and postmortems when capacity misallocation affects SLAs.
  • Maintain Jira epics and Confluence documentation for capacity planning, reporting, and change management.

Requirements

  • 5+ years of experience in technical program management, TPM, or product operations within cloud infrastructure, large-scale ML serving, or hyperscaler capacity planning.
  • Experience leading large cross-functional programs involving Engineering, Product, and Operations.
  • Understanding of inference serving concepts including model replicas, batching, prefill/decode, KV cache, and accelerator scheduling.
  • Strong data fluency with SQL, Grafana, and basic Python or Flux for independent analysis.
  • Experience running recurring cross-functional planning rituals with senior engineers and leadership teams.
  • Direct experience with AI accelerator fleet operations such as Habana, TPU pods, Inferentia, or Trainium.

Culture & Benefits

  • Opportunity to build an AI platform beyond the constraints of GPUs.
  • Access to cutting-edge AI research, model releases, and one of the fastest AI supercomputers.
  • Startup vitality combined with job stability.
  • Non-corporate culture focused on individual beliefs, continuous learning, growth, and inclusion.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →