Назад
Company hidden
8 часов назад

Hardware Production / Sustaining Engineer (AI/HPC)

170 000 - 205 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Hardware Production / Sustaining Engineer (AI/HPC): Driving the full lifecycle of GPU- and CPU-based compute hardware from prototype bring-up through high-volume production with an accent on PCIe, InfiniBand, NVMe/storage, and system reliability. Focus on designing hardware test automation, solving deep platform-level failures, and validating scalable AI/HPC infrastructure across multidisciplinary teams.

Location: Sunnyvale, California, United States; on-site

Salary: $170,000–$205,000 per year, plus Restricted Stock Units

Company

hirify.global builds vertically integrated, energy-focused AI infrastructure spanning energy, data centers, hardware, and cloud services.

What you will do

  • Own the hardware lifecycle from feasibility and prototype bring-up through validation, deployment, and production support.
  • Develop automation frameworks for hardware testing, diagnostics, and continuous reliability improvement using Python, Shell, or similar tools.
  • Debug PCIe link training, topology, and performance issues; InfiniBand fabric, throughput, and connectivity issues; and NVMe/storage bottlenecks, firmware interactions, and failures.
  • Conduct system validation and characterization for GPU, CPU, and high-performance computing platforms.
  • Lead end-to-end integration and solution testing for performance, reliability, and scalability.
  • Collaborate with mechanical, thermal, firmware, software, manufacturing teams, and external vendors to resolve system-level issues and prepare platforms for high-volume manufacturing.

Requirements

  • 4–8 years of experience in hardware development, validation, sustaining engineering, or production engineering.
  • Strong hands-on expertise in PCIe, InfiniBand, and NVMe/storage debugging and development.
  • Deep proficiency in hardware bring-up, board-level debugging, and system-level validation.
  • Technical background in digital and analog design, server architecture, and high-performance computing hardware.
  • Experience working across thermal, mechanical, firmware, and software functions in multidisciplinary environments.
  • Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, or equivalent experience.

Nice to have

  • Experience with GPU-to-GPU communication architectures, NVLink, advanced GPU architectures, and AI/HPC environments.
  • Experience supporting ARM and x86 server architectures.
  • Background in sustainable or energy-efficient hardware design.
  • Advanced certifications or coursework in AI/HPC hardware systems.

Culture & Benefits

  • Work on AI-first compute systems focused on performance, reliability, and sustainability.
  • Collaborate with experts across energy, manufacturing, data center construction, and cloud services.
  • Restricted Stock Units are included in all offers.
  • Equal opportunity employment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →