Назад
2 дня назад

Data Center Global Repairs Program Support

320 000 - 405 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Center Global Repairs Program Support (Data Center Hardware Operations): Define and manage global hardware repair programs across Anthropic’s data center fleet, covering server, GPU/accelerator, network, optics, RMA, reverse logistics, and repair inventory with an accent on repair SLAs, turnaround time, and production availability. Focus on standardizing triage and return-to-service procedures, analyzing failure patterns, managing vendors and OEMs, and driving corrective actions across complex HPC environments.

Location: Remote-friendly within the United States; staff are expected to work from one of the offices at least 25% of the time.

Annual salary: $320,000–$405,000 USD

Company

Anthropic develops reliable, interpretable, and steerable AI systems intended to be safe and beneficial for users and society.

What you will do

  • Define the global repair strategy, including SLAs, prioritization rules, escalation paths, reporting, and standardization across data center sites.
  • Own repair turnaround time, backlog, and compute returned to service using ticket and telemetry dashboards.
  • Develop triage, break-fix, and return-to-service procedures and train site operations partners.
  • Manage RMA, warranty, reverse logistics, and depot repair programs with OEMs, ODMs, and repair vendors.
  • Set spares pool sizing and stocking levels in coordination with supply chain and asset management.
  • Analyze failure trends, lead vendor and site operating reviews, and drive corrective actions with engineering and suppliers.

Requirements

  • 8+ years of data center operations experience as a manager, technical lead, or in a related role, including accountability for production availability.
  • Proven experience running large-scale break-fix programs across multiple sites.
  • Experience managing vendors, OEMs, ODMs, or contract workforces against measurable SLAs and corrective actions.
  • Hands-on technical depth in server, network, and rack-level hardware sufficient to verify repair quality and audit vendor claims.
  • Experience building or substantially improving operational processes and using ticket, telemetry, and inventory data for decisions.
  • Bachelor’s degree in a relevant field or equivalent practical experience.

Nice to have

  • Experience with GPU, accelerator, or high-density liquid-cooled infrastructure, including tray, cold plate, and manifold-level repair.
  • Experience with hyperscale OEM/ODM RMA, warranty, failure analysis, and supplier quality programs.
  • Experience with spares planning, reverse logistics, depot repair, partner-operated sites, colocation environments, optics, or high-speed interconnect failures.

Culture & Benefits

  • Flexible working hours and a collaborative office environment.
  • Competitive compensation with optional equity donation matching.
  • Generous vacation and parental leave.
  • Visa sponsorship is available, subject to role and candidate eligibility.
  • Anthropic is headquartered in San Francisco and operates as a public benefit corporation focused on safe and beneficial AI.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →