14 дней назад
Director of Product, Ecosystem (AI)
243 000 - 297 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Director of Product, Ecosystem (AI) (AI infrastructure and accelerator ecosystem): Defining ecosystem strategy for RDU-based token factories across networking, inference serving, storage, prefill compute, CPU offload, and compiler toolchains with an accent on partner integrations, reference architectures, and inference economics. Focus on converting co-optimization and certification work into deployed partner solutions, measurable demand, and partner-attached pipeline.
Location: San Jose, California, United States
Salary: $243,000–$297,000 USD base salary, plus equity and benefits.
Company
develops full-stack generative AI platforms, RDU inference accelerators, and rack-scale, managed-capacity, and API-based infrastructure for enterprise and government token factories.
What you will do
- Define ecosystem strategy for RDU integrations across scale-out networking, inference serving, storage and caching, prefill compute, agentic CPU offload, and compiler tooling.
- Decide whether integration areas should be built internally, delivered through partners, or deprioritized, using benchmarks, competitive intelligence, and unit economics.
- Publish joint reference architectures and own RDU certification across Kubernetes, Prometheus, Grafana, and managed-stack vendors.
- Drive co-optimization with vLLM, SGLang, llm-d, model partners, PyTorch, and MLIR communities.
- Lead joint go-to-market with token-factory partners and provide benchmark data for analyst relationships and buyer TCO models.
- Report quarterly on partner deployments, reference architectures, benchmark results, and partner-attached pipeline.
Requirements
- 10+ years in AI infrastructure, accelerators, or data center systems, including 8+ years in product management with direct ecosystem, platform, or partner-facing ownership.
- Experience shipping reference architectures, certification programs, or integration surfaces adopted by external teams.
- Working fluency across scale-out networking, inference serving, storage and caching, prefill and decode disaggregation, CPU offload, and compiler toolchains, with depth in at least two areas.
- Track record of converting partnerships into measurable deployments, demand, or revenue.
- Fluency in inference economics, including cost per million tokens, tokens per second per user, latency, and throughput tradeoffs.
- Bachelor's degree in computer science, electrical engineering, or a related technical field, or equivalent experience.
Nice to have
- Experience at both a silicon or systems vendor and a model lab or inference provider.
- Completed co-optimization engagements with model providers.
- Contributions to or close working relationships with vLLM, SGLang, llm-d, PyTorch, or MLIR communities.
- Exposure to hardware sales cycles, data center power and cooling constraints, and financing-backed deals.
- Experience running joint go-to-market with networking, storage, or compute vendors sharing a common buyer.
Culture & Benefits
- Full-time US employment with equity and a comprehensive benefits package.
- Medical insurance with 95% employee premium coverage and 77% dependent premium coverage.
- Health Savings Account with employer contribution, Dental, Vision, Disability, Life, AD&D, and Flexible Spending Account options.
- Well-being benefits including Headspace, Gympass+, One Medical, and counseling through an Employee Assistance Program.
- Equal opportunity employment for qualified applicants.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →