3 дня назад
Performance Analysis Engineer (ML Accelerator)
150 000 - 275 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Performance Analysis Engineer (ML Accelerator): Design and develop a sophisticated performance analysis tool tailored for custom ML accelerator hardware with an accent on data collection, analysis, and visualization of hardware and system-level performance metrics. Focus on building low-overhead tracing, correlating performance events across multiple hardware components, and enabling performance bottleneck identification and optimization.
Location
Location: San Jose, on-site
Salary
$150K – $275K per year plus significant equity
Company
builds hardware for frontier intelligence, co-designing chips, racks, software, and manufacturing to deliver best-in-class throughput and latency focused on inference workloads.
What you will do
- Lead design and architecture of a performance analysis suite including data collection, processing pipelines, analysis engines, and user interfaces.
- Develop methods to capture performance data from custom ML accelerator hardware via drivers or other mechanisms.
- Implement tracing for host-side API calls and system-level events related to workloads.
- Design techniques to correlate performance events across CPU, device drivers, PCIe bus, multiple accelerators, and hosts with precise time synchronization.
- Build analysis modules to identify key performance bottlenecks and develop intuitive visualizations to communicate findings.
- Collaborate closely with hardware architects, firmware, driver, compiler, and ML engineers to define tool requirements and provide expert guidance.
Requirements
- Location: Must work on-site in San Jose
- Strong proficiency in C++ or Rust; Python is a plus.
- Deep understanding of computer architecture, memory hierarchies, and PCIe interconnects.
- Proven experience in low-level performance analysis and profiling on complex hardware systems.
- Experience with performance analysis tools like NVIDIA Nsight, Intel VTune, perf, or similar.
- Experience working close to hardware, including reading performance counters or interacting with device drivers.
Nice to have
- Experience developing performance analysis or debugging tools.
- Experience with ML accelerator architectures (GPUs, TPUs).
- Kernel-mode driver development experience (Linux or Windows).
- Knowledge of compiler internals and optimization.
- In-depth knowledge of PCIe protocol and analysis tools.
- Experience with multi-chip or multi-host accelerator systems.
- Firmware or embedded systems development experience.
- Experience with hardware description languages or hardware verification.
Culture & Benefits
- Medical, dental, and vision insurance with generous coverage and $500/month credit for waiving medical benefits.
- Housing subsidy of $2,000/month for those living near the office.
- Relocation support for candidates moving to San Jose (Santana Row).
- Wellness benefits including fitness and mental health support.
- Daily lunch and dinner provided in the office.
- Unlimited compute budget subject to ROI justification.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
MTS - Kernel Engineer (AI)
260 000 - 320 000$
4 дня назад
MTS - Compiler Engineer (AI Hardware)
180 000 - 320 000$
3 дня назад
Senior Staff Software Engineer, Developer and Qualification Tools (AI)
175 000 - 265 000$
3 дня назад
Software Engineer (AI Hardware Systems)
3 дня назад
Senior Software Engineer (Performance)
193 930 - 291 150$
3 дня назад
Embedded Systems Engineer, Humanoid Robotics
135 000 - 185 000$