Назад
Company hidden
3 дня назад

Performance Analysis Engineer (ML Accelerator)

150 000 - 275 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Релокация
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Performance Analysis Engineer (ML Accelerator): Design and develop a sophisticated performance analysis tool tailored for custom ML accelerator hardware with an accent on data collection, analysis, and visualization of hardware and system-level performance metrics. Focus on building low-overhead tracing, correlating performance events across multiple hardware components, and enabling performance bottleneck identification and optimization.

Location

Location: San Jose, on-site

Salary

$150K – $275K per year plus significant equity

Company

hirify.global builds hardware for frontier intelligence, co-designing chips, racks, software, and manufacturing to deliver best-in-class throughput and latency focused on inference workloads.

What you will do

  • Lead design and architecture of a performance analysis suite including data collection, processing pipelines, analysis engines, and user interfaces.
  • Develop methods to capture performance data from custom ML accelerator hardware via drivers or other mechanisms.
  • Implement tracing for host-side API calls and system-level events related to workloads.
  • Design techniques to correlate performance events across CPU, device drivers, PCIe bus, multiple accelerators, and hosts with precise time synchronization.
  • Build analysis modules to identify key performance bottlenecks and develop intuitive visualizations to communicate findings.
  • Collaborate closely with hardware architects, firmware, driver, compiler, and ML engineers to define tool requirements and provide expert guidance.

Requirements

  • Location: Must work on-site in San Jose
  • Strong proficiency in C++ or Rust; Python is a plus.
  • Deep understanding of computer architecture, memory hierarchies, and PCIe interconnects.
  • Proven experience in low-level performance analysis and profiling on complex hardware systems.
  • Experience with performance analysis tools like NVIDIA Nsight, Intel VTune, perf, or similar.
  • Experience working close to hardware, including reading performance counters or interacting with device drivers.

Nice to have

  • Experience developing performance analysis or debugging tools.
  • Experience with ML accelerator architectures (GPUs, TPUs).
  • Kernel-mode driver development experience (Linux or Windows).
  • Knowledge of compiler internals and optimization.
  • In-depth knowledge of PCIe protocol and analysis tools.
  • Experience with multi-chip or multi-host accelerator systems.
  • Firmware or embedded systems development experience.
  • Experience with hardware description languages or hardware verification.

Culture & Benefits

  • Medical, dental, and vision insurance with generous coverage and $500/month credit for waiving medical benefits.
  • Housing subsidy of $2,000/month for those living near the office.
  • Relocation support for candidates moving to San Jose (Santana Row).
  • Wellness benefits including fitness and mental health support.
  • Daily lunch and dinner provided in the office.
  • Unlimited compute budget subject to ROI justification.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →