6 дней назад
Member Of Technical Staff, Infrastructure (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Member Of Technical Staff, Infrastructure (AI): Building and maintaining inference engine layers and benchmark suites for general-purpose AI systems across data center and on-device hardware with an accent on inference performance, model quality, correctness, and reproducibility. Focus on porting models across runtimes, designing explainable evaluations, and solving performance and memory tradeoffs across llama.cpp, ONNX, and MLX.
Location: Boston, United States; hybrid
Company
, spun out of MIT CSAIL, builds general-purpose AI systems optimized for efficient deployment across data center accelerators and on-device hardware.
What you will do
- Design and build benchmark suites covering inference performance, model quality, and knowledge evaluation across hardware targets.
- Run technical verifications for external partners, compare solutions against benchmarks, identify gaps, and communicate findings.
- Port models such as LFM2 across runtimes and frameworks and verify numerical correctness end to end.
- Maintain and extend the inference engine layer built on llama.cpp, ONNX, and MLX as new model architectures emerge.
- Make benchmark results explainable, verifiable, and reproducible for internal teams and external partners.
Requirements
- Hands-on experience with at least one inference framework such as llama.cpp, ONNX Runtime, or MLX, including internals and modification.
- Experience designing and building benchmarking pipelines with rigorous methodology, validation, and reproducibility.
- Strong C++ and Python skills in performance-sensitive contexts.
- Solid understanding of inference fundamentals, including quantization, decoding strategies, memory layout, and their tradeoffs.
Nice to have
- Experience porting models across runtimes and verifying numerical correctness.
- Experience working with external partners or clients in technical validation or evaluation.
- Familiarity with edge inference targets and their constraints.
Culture & Benefits
- Work closely with research, product, and external engineering teams.
- Medical, dental, and vision premiums fully covered for employees and dependents.
- 401(k) matching of up to 4% of base pay.
- Unlimited paid time off and company-wide Refill Days.
- Competitive base salary and equity participation.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
16 часов назад
Member of Technical Staff — Model Optimization and Inference (Experienced)
250 000 - 350 000$
6 дней назад
Staff Software Engineer (AI Foundation Model)
5 дней назад
AI Infrastructure Engineer
170 500 - 315 490$
2 дня назад
Member of Technical Staff (AI)
100 000 - 140 000€
3 дня назад
Forward Deployed Engineer (AI)
3 дня назад