Назад
Company hidden
14 дней назад

Director of Product, Ecosystem (AI)

243 000 - 297 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
director
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Director of Product, Ecosystem (AI) (AI infrastructure and accelerator ecosystem): Defining ecosystem strategy for RDU-based token factories across networking, inference serving, storage, prefill compute, CPU offload, and compiler toolchains with an accent on partner integrations, reference architectures, and inference economics. Focus on converting co-optimization and certification work into deployed partner solutions, measurable demand, and partner-attached pipeline.

Location: San Jose, California, United States

Salary: $243,000–$297,000 USD base salary, plus equity and benefits.

Company

hirify.global develops full-stack generative AI platforms, RDU inference accelerators, and rack-scale, managed-capacity, and API-based infrastructure for enterprise and government token factories.

What you will do

  • Define ecosystem strategy for RDU integrations across scale-out networking, inference serving, storage and caching, prefill compute, agentic CPU offload, and compiler tooling.
  • Decide whether integration areas should be built internally, delivered through partners, or deprioritized, using benchmarks, competitive intelligence, and unit economics.
  • Publish joint reference architectures and own RDU certification across Kubernetes, Prometheus, Grafana, and managed-stack vendors.
  • Drive co-optimization with vLLM, SGLang, llm-d, model partners, PyTorch, and MLIR communities.
  • Lead joint go-to-market with token-factory partners and provide benchmark data for analyst relationships and buyer TCO models.
  • Report quarterly on partner deployments, reference architectures, benchmark results, and partner-attached pipeline.

Requirements

  • 10+ years in AI infrastructure, accelerators, or data center systems, including 8+ years in product management with direct ecosystem, platform, or partner-facing ownership.
  • Experience shipping reference architectures, certification programs, or integration surfaces adopted by external teams.
  • Working fluency across scale-out networking, inference serving, storage and caching, prefill and decode disaggregation, CPU offload, and compiler toolchains, with depth in at least two areas.
  • Track record of converting partnerships into measurable deployments, demand, or revenue.
  • Fluency in inference economics, including cost per million tokens, tokens per second per user, latency, and throughput tradeoffs.
  • Bachelor's degree in computer science, electrical engineering, or a related technical field, or equivalent experience.

Nice to have

  • Experience at both a silicon or systems vendor and a model lab or inference provider.
  • Completed co-optimization engagements with model providers.
  • Contributions to or close working relationships with vLLM, SGLang, llm-d, PyTorch, or MLIR communities.
  • Exposure to hardware sales cycles, data center power and cooling constraints, and financing-backed deals.
  • Experience running joint go-to-market with networking, storage, or compute vendors sharing a common buyer.

Culture & Benefits

  • Full-time US employment with equity and a comprehensive benefits package.
  • Medical insurance with 95% employee premium coverage and 77% dependent premium coverage.
  • Health Savings Account with employer contribution, Dental, Vision, Disability, Life, AD&D, and Flexible Spending Account options.
  • Well-being benefits including Headspace, Gympass+, One Medical, and counseling through an Employee Assistance Program.
  • Equal opportunity employment for qualified applicants.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →