9 часов назад
Senior Software Engineering Manager (AI Infrastructure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Software Engineering Manager (AI Infrastructure): Leading the engineering organization building a distributed KV Cache Platform that accelerates large-scale LLM inference across GPU clusters with an accent on scalable distributed systems, low-latency storage, and GPU infrastructure. Focus on defining the technical roadmap, driving architecture and production delivery, and building high-performing geographically distributed engineering teams.
Location: Santa Clara, United States; hybrid
Company
develops AI infrastructure and distributed memory and storage platforms for large-scale LLM inference across GPU clusters.
What you will do
- Lead, mentor, and grow geographically distributed software engineering teams and technical leaders.
- Define and execute the technical strategy and roadmap for the KV Cache Platform.
- Drive the architecture, development, and delivery of distributed systems for AI inference, GPU memory optimization, distributed caching, RDMA networking, GPUDirect Storage, and NVIDIA BlueField DPUs.
- Own engineering execution, including feature development, release planning, bug triage, production issues, customer escalations, and cross-functional delivery.
- Partner with Product Management, Sales, Customer Engineering, NVIDIA, strategic technology partners, and executive leadership on customer requirements, POCs, product direction, and deployments.
- Establish practices for software quality, observability, automation, performance, testing, and production readiness.
Requirements
- 15+ years of experience building distributed systems, cloud infrastructure, storage platforms, or AI infrastructure software.
- 7+ years leading high-performing software engineering organizations, including geographically distributed teams.
- Experience delivering large-scale distributed infrastructure products from architecture through production deployment.
- Strong background in distributed systems, Linux, networking, performance engineering, and cloud-native architectures.
- Hands-on programming experience with Go and Python.
- Experience with AI infrastructure, LLM serving platforms, distributed caching, high-performance storage, HPC, networking, or enterprise infrastructure software.
Nice to have
- Experience with C or C++.
- Experience with NVIDIA Dynamo, TensorRT-LLM, Triton, RDMA, GPUDirect Storage, BlueField DPUs, Kubernetes, or related AI infrastructure.
- Experience working with strategic customers, technology partners, OEMs, or hyperscalers to deliver enterprise AI solutions.
Culture & Benefits
- Technical leadership role focused on engineering excellence, innovation, ownership, and collaboration.
- Opportunity to shape AI infrastructure product strategy and support customer deployments.
- Collaboration across engineering, infrastructure, hardware, product, sales, customer engineering, and technology partner organizations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Director of Engineering (AI)
225 000 - 280 000$
10 часов назад
Engineering Manager (AI Infrastructure)
250 000 - 340 000$
8 часов назад
Senior Engineering Manager (AI)
215 000 - 250 000$
12 часов назад
Director of Engineering (AI)
250 000 - 330 000$
2 дня назад
Support Engineering Manager (Go/Python)
137 000 - 172 000$
5 часов назад