10 дней назад
Compute Lead (AI)
200 000 - 400 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Compute Lead (AI infrastructure): Managing compute allocation, capacity sourcing, cluster acceptance, and GPU utilization for model training, research, and customer workloads with an accent on infrastructure operations, vendor negotiations, and high-stakes resource allocation. Focus on forecasting compute demand, testing and monitoring clusters, and ensuring research and product roadmaps have sufficient infrastructure.
Location: San Francisco, CA; onsite five days a week
Salary: $200,000–$400,000 per year, plus equity and competitive benefits.
Company
is an AI interpretability research lab building interpretability agents and infrastructure to understand, monitor, and align AI models.
What you will do
- Manage compute allocations across model training, research, and customer workloads.
- Source and secure compute capacity by comparing provider availability and pricing and managing deals through contracting and delivery.
- Define cluster acceptance criteria and run burn-in and performance testing with the infrastructure team.
- Monitor live cluster health, node failures, downtime, utilization, and quota usage, driving fixes with providers.
- Partner with research and product teams to translate roadmaps into compute requirements and prevent infrastructure bottlenecks.
- Build forecasting models for compute demand, long-term planning, and budgeting.
Requirements
- Experience navigating GPU scarcity, provider negotiations, and compute capacity decisions.
- Strong judgment when balancing competing demands and making high-stakes resource allocation trade-offs.
- Ability to work across infrastructure, operations, research, and product functions.
- Willingness to work onsite in the San Francisco office five days per week.
- Ability to operate with urgency, ownership, initiative, and a focus on practical improvements.
Culture & Benefits
- Mission focused on making AI interpretability accessible and advancing safe, beneficial AI.
- High-intensity environment with an emphasis on taking action quickly.
- Values include putting the mission and team first, continuous improvement, and personal ownership.
- Market-competitive salary, equity, and competitive benefits.
- Opportunity to work with a world-class team on AI interpretability research and infrastructure.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
Senior/Staff DevOps Engineer (AI Platform Infrastructure)
180 000 - 240 000$
13 дней назад
Engineering Manager, Platform & Infrastructure (AI)
200 000 - 300 000$
12 дней назад
Tech Lead (AWS/TypeScript)
200 000 - 270 000$
13 дней назад
Forward Deployed Engineer (AI/Docker)
198 000 - 319 000$
10 дней назад
AI / DevOps Engineer (Agentic Systems & Automation)
139 000 - 243 000$
13 дней назад
Principal Engineering Lead
174 300 - 205 900$