7 часов назад
Member of Technical Staff (AI Infrastructure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Member of Technical Staff (AI Infrastructure): Building production systems, research prototypes, and experimental platforms for power-aware AI infrastructure and large-scale cloud computing with an accent on distributed systems, machine learning, optimization, and GPU operations. Focus on designing algorithms, deploying research innovations into AI data centers, and solving complex resource scheduling, performance, and energy-management challenges across thousands of GPUs.
Location: Hybrid in the Bay Area, Boston, or Washington D.C.; 2 work-from-home days per week
Company
develops software that helps AI data centers dynamically manage power usage, enabling sustainable AI growth while supporting grid reliability and renewable energy expansion.
What you will do
- Design and build systems for power-aware AI infrastructure, distributed computing, and large-scale cloud platforms.
- Develop production-quality software, research prototypes, and experimental infrastructure for real-world AI data centers.
- Apply machine learning, optimization, systems, and control techniques to AI infrastructure and cloud operations.
- Design, implement, and evaluate algorithms using large-scale experimental platforms and production deployments.
- Translate research innovations into customer-facing products in collaboration with product and customer teams.
- Collaborate with industry and academic partners, publish research when appropriate, and help shape the technical roadmap.
Requirements
- Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a closely related field.
- Strong background in machine learning systems, AI infrastructure, distributed systems, cloud computing, systems for AI or HPC, or performance optimization.
- Excellent software engineering skills with C++, Python, Go, or Rust.
- Experience building research prototypes or large-scale production software and independently driving research through implementation and evaluation.
- Strong publication record or a demonstrated history of delivering impactful technical innovations.
- Ability to work in a hybrid setup from the Bay Area, Boston, or Washington D.C.
Nice to have
- Experience with Kubernetes, Slurm, distributed training or inference frameworks, or large-scale AI infrastructure.
- Experience with GPU systems, accelerators, or performance analysis tools.
- Experience in optimization, control systems, resource scheduling, systems performance, power management, energy-efficient computing, sustainability, or data center infrastructure.
- Experience deploying systems in production cloud or distributed environments and taking research innovations from prototype to production.
Culture & Benefits
- Collaborative, low-ego environment with experts in AI, cloud, software, and energy.
- Opportunity to build new products and influence strategy, go-to-market, organizational design, and customer engagement.
- Competitive compensation with stock options.
- Medical, dental, and vision insurance, plus 401(k) matching.
- Equal opportunity workplace with reasonable accommodations available throughout the interview process.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
7 часов назад
Member of Technical Staff (Applied AI)
150 000 - 350 000$
7 часов назад
R&D Software Engineer (AI)
6 часов назад
AI Engineer
11 часов назад
Member of Technical Staff - Agent Engineer (AI)
200 000 - 300 000$
2 дня назад
Staff Software Engineer (Agentic AI)
7 часов назад
Member of Technical Staff, Training Infra (AI)
200 000 - 350 000$