4 часа назад
Member of Technical Staff, Inference (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Member of Technical Staff, Inference (AI): Building and optimizing production inference systems for large-scale genome and multimodal biological AI models with an accent on GPU kernels, distributed computing, and performance engineering. Focus on reducing latency for large MoE models, maximizing throughput across heterogeneous hardware, and deploying foundation models reliably across cloud, enterprise, and secure environments.
Location: San Francisco; authorization to work in the United States is required
Company
is an AI research lab developing generative biological intelligence and foundation models for genomics, therapeutics, diagnostics, synthetic biology, and biodefense.
What you will do
- Drive end-to-end performance improvements across model execution, memory management, networking, scheduling, and hardware utilization.
- Develop and optimize GPU kernels, numerical operators, and inference-serving infrastructure.
- Build scalable systems for serving, monitoring, benchmarking, and operating foundation models across cloud, enterprise, and secure environments.
- Deploy biological AI models with pharmaceutical, biotech, research, and government partners.
- Collaborate with research, infrastructure, and platform teams to bring new model architectures into production.
Requirements
- Production experience optimizing, deploying, and operating large language models or other foundation models.
- Deep understanding of GPU architectures and hands-on experience with CUDA, Triton, or equivalent technologies.
- Ability to diagnose bottlenecks across model architectures, serving systems, networking, memory management, and distributed infrastructure.
- Strong software engineering skills with proficiency in Python and PyTorch or comparable modern ML frameworks.
- Experience working with customers, users, or cross-functional stakeholders to deliver production AI systems.
- Must be authorized to work in the United States; participates in E-Verify.
Nice to have
- Experience with vLLM, TensorRT-LLM, SGLang, DeepSpeed, or similar inference frameworks.
- Contributions to open-source AI infrastructure, inference frameworks, compilers, or kernel libraries.
- Experience with distributed systems, cloud infrastructure, and large-scale GPU clusters.
- Background in biological foundation models, computational biology, genomics, or scientific AI.
- Experience operating AI systems in regulated, secure, or mission-critical environments.
Culture & Benefits
- Work on large open biological AI models with applications in therapeutics, diagnostics, synthetic biology, and biodefense.
- Collaborate with researchers across AI labs, biotechs, pharmaceutical companies, hospitals, government programs, and scientific institutions.
- Join a multidisciplinary environment spanning distributed systems, model architecture, numerics, AI safety, and biology.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 часа назад
Member of Technical Staff — Inference-Multi-Hardware (AI Infrastructure)
200 000 - 400 000$
4 часа назад
Member of Technical Staff — Inference-Core Engine (AI)
200 000 - 400 000$
4 часа назад
Member of Technical Staff, Kernels (AI)
200 000 - 350 000$
4 часа назад
Member of Technical Staff — Inference-Kernel, Compiler & Communication (AI)
200 000 - 400 000$
4 часа назад
Member of Technical Staff, Inference & Serving (AI)
200 000 - 350 000$
5 часов назад
Member of Technical Staff (Research Infrastructure)
200 000 - 400 000$