6 дней назад
Superintelligence Technical Success Engineer (AI)
251 000 - 335 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Superintelligence Technical Success Engineer (GPU/HPC Infrastructure): Taking signed AI infrastructure deployments from contract through validation, onboarding, and live production with an accent on GPU/HPC systems, cloud platforms, Kubernetes, Linux, networking, storage, and compute. Focus on troubleshooting deployment gaps, coordinating engineering and infrastructure teams, managing technical risks, and guiding customers to their first successful production workload.
Location: Hybrid, based in the San Francisco or San Jose office 4 days per week; Tuesday is the designated work-from-home day.
Salary: $251K–$335K annual salary.
Company
Lambda provides AI cloud infrastructure for AI researchers, enterprises, and hyperscalers.
What you will do
- Take signed deployments from contract through configuration, connectivity, storage, and compute validation to live production environments.
- Troubleshoot technical issues, identify gaps, and work with engineering and infrastructure teams to resolve blockers.
- Coordinate with Infrastructure, Engineering, Product, and Data Center teams on technical dependencies.
- Maintain an accurate view of deployment progress, open issues, risks, and mitigations.
- Guide customers through onboarding to their first successful production workload and act as the primary technical contact during deployment and early production.
- Turn recurring technical patterns into reusable runbooks, checklists, or automation.
Requirements
- 4+ years of hands-on technical experience with GPU/HPC infrastructure, cloud platforms, Kubernetes, or large-scale Linux systems.
- Ability to validate technical builds, identify gaps, and drive engineering teams toward resolution.
- Strong troubleshooting skills across networking, storage, and compute.
- Experience coordinating engineering and infrastructure teams to resolve technical dependencies.
- Clear written and verbal communication for status updates and technical documentation.
- Ability to work from the San Francisco or San Jose office 4 days per week.
Nice to have
- Experience with large-scale GPU cluster deployments.
- Familiarity with project or program tracking tools and structured status reporting.
- Experience creating durable runbooks or checklists.
Culture & Benefits
- Cash and equity compensation.
- Health, dental, and vision coverage for employees and dependents.
- Wellness and commuter stipends for select roles.
- 401(k) plan with a 2% company match for USA employees.
- Flexible paid time off.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Tactical Systems Engineer (C/C++/Python)
114 000 - 231 000$
4 дня назад
Hardware Systems Engineer (AI/HPC)
92 000 - 112 000$
7 дней назад
Implementation Engineer
140 000 - 200 000$
6 дней назад
High Performance Compute Responsible Engineer (Aerospace)
175 000 - 263 000$
5 дней назад
Firmware/Software Tech Lead (AI Infrastructure)
160 000 - 200 000$
5 часов назад
Lead Rack Systems Engineer (AI)
208 000 - 263 000$