3 часа назад
Senior AI Systems Performance Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior AI Systems Performance Engineer (AI) (Foundation Models and ML Systems): Optimizing and scaling foundation models on SambaNova’s reconfigurable dataflow platform with an accent on throughput, latency, compiler-runtime-hardware collaboration, and efficient inference. Focus on profiling bottlenecks, applying quantization and scheduling optimizations, and building scalable single-node and distributed AI systems.
Location: San Jose, California, United States
Company
develops a full-stack generative AI platform spanning specialized chips, software, and foundation models for enterprise and government organizations.
What you will do
- Bring up and optimize foundation models such as DeepSeek, Llama, and Qwen on the software stack.
- Profile and improve model performance across compiler, runtime, and hardware layers.
- Collaborate with machine learning, compiler, runtime, and hardware teams on co-designed AI applications.
- Apply advances in model architecture, quantization, scheduling, and memory optimization.
- Develop scalable end-to-end inference solutions for customer needs.
- Identify bottlenecks and optimize dataflow and scheduling for single-node and distributed systems.
Requirements
- Bachelor’s degree or higher in computer science, electrical engineering, applied mathematics, physics, statistics, or a related field.
- 3+ years of experience in deep learning performance optimization, compiler/runtime/kernel optimization, systems tuning, or software-hardware co-design.
- Proficiency in Python or C++, with strong foundations in algorithms, data structures, and numerical computing.
- Experience with at least one major ML framework: PyTorch, TensorFlow, or JAX.
- Demonstrated ability to analyze and optimize real-world ML pipelines.
Nice to have
- Experience with LLM or multimodal model training and inference.
- Background in distributed training, continuous batching, and high-throughput inference.
- Familiarity with quantization, graph optimization, kernel fusion, model partitioning, DeepSpeed, Megatron, vLLM, or TensorRT.
- GPU programming experience with CUDA, Triton, or OpenCL, and familiarity with cuDNN or cuBLAS.
- Publications or open-source contributions in ML systems or performance optimization.
Culture & Benefits
- Full-time US employment includes base salary, equity, and benefits.
- Medical insurance covers 95% of employee premiums and 77% of dependent premiums.
- Health Savings Account, Flexible Spending Account, dental, vision, disability, life, and AD&D insurance options are available.
- Well-being benefits include Headspace, Gympass+, One Medical, and counseling services.
- is an Equal Opportunity/Affirmative Action Employer.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 час назад
AI Performance Engineer
Top Selection
5 дней назад
Senior AI Developer (Python/C++)
1 минуту назад
AI Systems Engineer (AI)
150 000 - 225 000$
Samsara
5 дней назад
Senior Software Engineer (AI Platform)
2 дня назад
Senior Applied Scientist (AI)
183 800 - 248 700$
1 час назад
Software Engineer (AI)
145 000 - 155 000$