8 часов назад
Hardware Production / Sustaining Engineer (AI/HPC)
170 000 - 205 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Hardware Production / Sustaining Engineer (AI/HPC): Driving the full lifecycle of GPU- and CPU-based compute hardware from prototype bring-up through high-volume production with an accent on PCIe, InfiniBand, NVMe/storage, and system reliability. Focus on designing hardware test automation, solving deep platform-level failures, and validating scalable AI/HPC infrastructure across multidisciplinary teams.
Location: Sunnyvale, California, United States; on-site
Salary: $170,000–$205,000 per year, plus Restricted Stock Units
Company
builds vertically integrated, energy-focused AI infrastructure spanning energy, data centers, hardware, and cloud services.
What you will do
- Own the hardware lifecycle from feasibility and prototype bring-up through validation, deployment, and production support.
- Develop automation frameworks for hardware testing, diagnostics, and continuous reliability improvement using Python, Shell, or similar tools.
- Debug PCIe link training, topology, and performance issues; InfiniBand fabric, throughput, and connectivity issues; and NVMe/storage bottlenecks, firmware interactions, and failures.
- Conduct system validation and characterization for GPU, CPU, and high-performance computing platforms.
- Lead end-to-end integration and solution testing for performance, reliability, and scalability.
- Collaborate with mechanical, thermal, firmware, software, manufacturing teams, and external vendors to resolve system-level issues and prepare platforms for high-volume manufacturing.
Requirements
- 4–8 years of experience in hardware development, validation, sustaining engineering, or production engineering.
- Strong hands-on expertise in PCIe, InfiniBand, and NVMe/storage debugging and development.
- Deep proficiency in hardware bring-up, board-level debugging, and system-level validation.
- Technical background in digital and analog design, server architecture, and high-performance computing hardware.
- Experience working across thermal, mechanical, firmware, and software functions in multidisciplinary environments.
- Bachelor’s or Master’s degree in Electrical Engineering, Computer Engineering, or equivalent experience.
Nice to have
- Experience with GPU-to-GPU communication architectures, NVLink, advanced GPU architectures, and AI/HPC environments.
- Experience supporting ARM and x86 server architectures.
- Background in sustainable or energy-efficient hardware design.
- Advanced certifications or coursework in AI/HPC hardware systems.
Culture & Benefits
- Work on AI-first compute systems focused on performance, reliability, and sustainability.
- Collaborate with experts across energy, manufacturing, data center construction, and cloud services.
- Restricted Stock Units are included in all offers.
- Equal opportunity employment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 часов назад
Post Silicon Validation Engineer (AI/ML)
160 000 - 260 000$
8 часов назад
Hardware Validation Engineer
150 000 - 275 000$
9 часов назад
Board Design Engineer
150 000 - 260 000$
8 часов назад
Electrical Engineer (AI Hardware)
190 000 - 230 000$
9 часов назад
DVT Technical Lead (AI Hardware)
175 000 - 275 000$
9 часов назад
Physical Design Engineer (AI Hardware)
150 000 - 270 000$