Назад
1 день назад

Lead Hardware Deployment Engineer (AI)

Формат работы
remote (только USA)/hybrid
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Lead Hardware Deployment Engineer (AI): Building and leading the end-to-end bring-up of GPU compute hardware for the world's largest AI training clusters with an accent on L11 rack integration and post-integration hardware health. Focus on optimizing deployment velocity, enforcing vendor SLAs, and maximizing node availability.

Location: Based in Memphis, TN, USA (On-site presence required)

Company

AI company focused on creating systems that accurately understand the universe and aid humanity in its pursuit of knowledge.

What you will do

  • Lead, hire, and develop a dedicated hardware deployment team of engineers and technicians.
  • Own L11 rack integration and compute hardware bring-up across multiple data halls concurrently.
  • Drive aggressive timelines to achieve 95%+ node availability within days of rack delivery.
  • Run systematic health pushes to sustain over 98% node availability prior to operations turnover.
  • Internalize non-RMA hardware repairs to maximize recovery and minimize OEM turnaround times.
  • Enforce vendor SLAs for OEMs and suppliers to prevent repair backlogs.

Requirements

  • 5+ years of experience deploying or repairing compute/server hardware at data center scale.
  • Direct experience with L11 rack-level integration of GPU or accelerator-based systems.
  • Experience leading technician or engineering teams in fast-paced deployment environments.
  • Deep troubleshooting skills in servers, GPUs, NVLink, high-speed networking, and liquid cooling.
  • Willingness to work on-site in Memphis, TN, including extended hours and weekends during critical phases.

Nice to have

  • Experience with NVIDIA GB200/GB300 NVL72 or similar rack-scale liquid-cooled systems.
  • Experience building a new team or function from scratch, including hiring and training.
  • Experience managing OEM/ODM vendor relationships (e.g., Dell, Supermicro).
  • Knowledge of data center automation, burn-in/validation tooling, and hardware health telemetry.

Culture & Benefits

  • Flat organizational structure where all employees are expected to be hands-on.
  • Environment that values engineering excellence, curiosity, and strong work ethic.
  • High-impact role contributing directly to the company's core mission on a critical path.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →