3 дня назад
Infrastructure Hardware Technical Program Manager (Server and Network Systems)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Infrastructure Hardware Technical Program Manager (Server and Network Systems) (AI infrastructure): Driving end-to-end delivery of server and network platform programs for Cerebras CS-3-based AI clusters with an accent on requirements, vendor coordination, qualification, and production readiness. Focus on aligning architects, OEM/ODM partners, validation teams, and deployment operations while managing technical risks, dependencies, and rollout decisions.
Location: US and Canada offices
Company
builds large-scale AI hardware and software platforms designed to deliver faster model training and inference than GPU-based systems.
What you will do
- Own end-to-end delivery of server and network platform programs for CS-3-based AI clusters, including new platforms, refreshes, and major configuration changes.
- Gather requirements and create integrated plans with milestones, readiness gates, dependencies, risks, and cross-functional deliverables.
- Manage OEM/ODM partners, switch vendors, component suppliers, RFIs/RFPs, samples, escalations, and roadmap alignment.
- Translate architectural decisions into qualification plans, acceptance criteria, release-readiness reviews, and rollout strategies.
- Lead lab and staging validation, regression tracking, go/no-go decisions, production change management, and rollout sequencing.
- Coordinate operational readiness with deployment and fleet teams while aligning server and network changes with rack and physical data center constraints.
Requirements
- Bachelor’s or master’s degree in Computer Science, Electrical Engineering, Computer Engineering, or equivalent experience.
- 8+ years of technical program management or delivery leadership experience for server, network, or infrastructure platforms from concept through production.
- Experience coordinating server or data center network programs across OEM/ODM partners, switch vendors, and internal engineering teams.
- Working knowledge of server architecture, including CPU/NUMA, memory bandwidth, PCIe, NICs, storage I/O, and networking fundamentals such as leaf-spine fabrics and high-performance interconnects.
- Familiarity with Linux server fleet management, provisioning, firmware/BIOS, drivers, and field triage.
- Strong multi-team execution, risk management, dependency tracking, ambiguity management, and executive communication skills.
Nice to have
- Experience with AI/ML, HPC, or performance-sensitive distributed infrastructure.
Culture & Benefits
- Work on a large-scale AI platform designed beyond traditional GPU constraints.
- Opportunities to publish and open-source AI research.
- Work with one of the fastest AI supercomputers in the world.
- Startup vitality combined with job stability.
- Simple, non-corporate culture with respect for individual beliefs, continuous learning, and professional growth.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Technical Program Manager, Infrastructure (AI)
4 дня назад
ASIC Technical Program Manager (AI)
3 дня назад
Staff Technical Program Manager (Cybersecurity)
178 000 - 213 000$
17 часов назад
Technical Program Manager
5 дней назад
Senior Technical Program Manager (AI/ML)
131 000 - 245 000$
3 дня назад
Engineering Program Manager (AI Infrastructure)
230 000 - 280 000$