6 часов назад
ML Performance Engineer
200 000 - 350 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
ML Performance Engineer (AI infrastructure): Improving the performance of Modal’s container runtime and open-source projects for language and diffusion model workloads with an accent on GPU throughput, inference latency, and scalable ML systems. Focus on optimizing CUDA workloads, diagnosing GPU performance bottlenecks, and reducing host overhead in production inference systems.
Location: New York or San Francisco, United States
Salary: $200,000–$350,000 annually, plus equity
Company
is building infrastructure for AI workloads, providing instant GPU access, fast container starts, native storage, production-ready sandboxes, and low-latency model inference at scale.
What you will do
- Improve the performance and scalability of ML systems.
- Contribute to open-source projects and ’s container runtime.
- Optimize language and diffusion model workloads for higher throughput and lower latency.
- Investigate GPU performance bottlenecks, including SM occupancy and host overhead.
- Rewrite algorithms and systems components to improve compute efficiency.
Requirements
- 5+ years of experience writing high-quality, high-performance code.
- Experience with Torch, high-level ML frameworks, and inference engines such as vLLM or TensorRT.
- Familiarity with Nvidia GPU architecture and CUDA.
- Experience with ML performance engineering and demonstrated GPU performance optimization results.
Nice to have
- Familiarity with low-level operating system foundations, including the Linux kernel, file systems, and containers.
Culture & Benefits
- Opportunity to contribute to open-source projects.
- Work alongside researchers, open-source creators, and experienced engineering and product leaders.
- Equity included in the compensation package.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 часов назад
Distributed LLM Inference Engineer (AI)
170 000 - 245 000$
6 часов назад
Customer Engineer (ML/AI)
170 000 - 199 000$
7 часов назад
Member of Technical Staff - ML Systems & Inference
250 000 - 350 000$
6 часов назад
Research Scientist (Embodied AI)
250 000 - 350 000$
5 часов назад
Staff Software Engineer (SIMD Kernels)
190 000 - 300 000$
5 часов назад
Principal Software Engineer, Kernels (AI)
195 000 - 285 000$