Назад
22 часа назад

Hardware Integration Engineer (AI)

255 000 - 340 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Hardware Integration Engineer (AI Cloud): Owning system integration validation for HPC AI/ML and compute platforms with an accent on NPI, reliability testing, and scale deployment. Focus on developing automated stress/benchmark tests, managing firmware compatibility, and performing root-cause analysis for fleet hardware issues.

Location: Hybrid: Must be based in San Jose, CA (office presence required 4 days per week)

Salary: $255K – $340K

Company

Lambda is a leader in AI cloud infrastructure, providing high-performance compute resources for AI researchers and enterprises.

What you will do

  • Own system integration validation for new HPC AI/ML, compute, storage, and network hardware platforms during NPI.
  • Develop and execute functional, stress, reliability, and benchmark tests, and enable automation for scale testing.
  • Collaborate with deployment and fleet engineering teams to facilitate L10/L11/L12 level benchmarking.
  • Manage firmware and software compatibility and support maintenance after production release.
  • Conduct root-cause analysis (RCA) for critical hardware issues and perform failure correlation and trend analysis.
  • Establish lab setup and test infrastructure for repeatable hardware evaluation in partnership with data center and network design teams.

Requirements

  • 5+ years of experience in hardware integration validation, fleet hardware reliability, or HPC performance validation.
  • Deep knowledge of system integration testing and performance benchmarking at L10, L11, and L12 levels.
  • Experience validating AI/ML, x86, ARM, storage systems, or network switches.
  • Hands-on experience with vendor-led NPI cycles, including debugging and triage at scale.
  • Ability to collaborate across architecture, supply chain, and infrastructure engineering teams.

Nice to have

  • 10+ years of experience in hardware integration and reliability engineering.
  • Experience supporting accelerated compute hardware from NVIDIA, AMD, or Intel.
  • Deep knowledge of BMC, BIOS settings, and network card configurations.
  • Proficiency in performance tuning for rack-scale servers.
  • Working knowledge of PLM systems and BOM structures.

Culture & Benefits

  • Competitive cash and equity compensation.
  • Comprehensive health, dental, and vision coverage for employees and dependents.
  • 401k plan with a 2% company match for US employees.
  • Wellness and commuter stipends for select roles.
  • Flexible paid time off policy.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →