Назад
7 часов назад

AI Infrastructure Operations, Demand Planning

320 000 - 405 000$
Формат работы
hybrid
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Infrastructure Operations, Demand Planning (AI infrastructure): Converts demand forecasts into tranche-level infrastructure requirements and manages capacity from contract through occupied production state with an accent on accelerator capacity, regional delivery, and infrastructure readiness. Focus on coordinating multi-region bring-ups, defining lifecycle systems of record, automating readiness, and reporting delivery risk and utilization economics across large-scale infrastructure fleets.

Location: San Francisco, CA or New York City, NY; staff are expected to work from one of the offices at least 25% of the time.

Annual salary: $320,000–$405,000 USD.

Company

Anthropic is a public benefit corporation building reliable, interpretable, and steerable AI systems.

What you will do

  • Translate demand forecasts and stakeholder input into accelerator, interconnect, regional, supporting-resource, and delivery-date requirements for each capacity tranche.
  • Assess whether proposed capacity can be delivered as schedulable, healthy, and instrumented infrastructure before contracts are signed.
  • Track forecast-versus-delivered variance and feed delivery findings into demand planning, sourcing negotiations, and future contracts.
  • Define and maintain the contract-to-occupied lifecycle state machine and canonical system of record for in-flight capacity.
  • Coordinate parallel cloud, on-premises, and neocloud bring-ups across provider milestones, cluster creation, networking, storage, health burn-in, and workload landing.
  • Automate readiness workflows and publish executive reporting on time-to-occupied, paid-idle costs, status, trade-offs, and portfolio risk.

Requirements

  • Significant experience delivering large-scale infrastructure such as cloud regions, accelerator clusters, HPC systems, or bare-metal fleets at multi-region scale or equivalent capacity.
  • Technical understanding spanning cluster orchestration, node health, telemetry, capacity data, and planning systems.
  • Experience using SQL and enough Python to investigate infrastructure data and build reporting.
  • Bachelor’s degree in a relevant technical field or equivalent engineering experience.
  • Ability to work from the San Francisco or New York City office at least 25% of the time.

Nice to have

  • Experience with reserved-capacity onboarding, private offers, or cloud and neocloud capacity commitments.
  • Demand-planning experience, including challenging forecasts and translating them into tranche-level requirements.
  • Data center or colocation delivery, power and space planning, network turn-up, site acceptance, or vendor management.
  • Experience with accelerator health, burn-in, collective-communications testing, fleet-health SLOs, or infrastructure lifecycle systems.
  • Experience onboarding new hardware generations into scheduler and observability systems.

Culture & Benefits

  • Collaborative, research-driven environment focused on large-scale efforts in safe and beneficial AI.
  • Frequent research discussions and strong emphasis on communication and cross-functional collaboration.
  • Flexible working hours and office collaboration.
  • Competitive compensation, optional equity donation matching, generous vacation, and parental leave.
  • Visa sponsorship may be available, with immigration-lawyer support, depending on the role and candidate.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →