Назад
Company hidden
4 дня назад

Product Manager (AI Inference Platform)

170 000 - 300 000$
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Product Manager (AI Inference Platform) (AI infrastructure): Owns strategy, roadmap, and success metrics for inference and platform products across API, UI, and CLI, with an accent on model serving, platform controls, accounts, billing, and fraud prevention. Focus on balancing latency, throughput, GPU utilization, cost, rate limits, and reliability while turning production customer pain into scalable product capabilities.

Location: San Mateo, United States

Salary: $170,000–$300,000 per year, plus equity.

Company

hirify.global provides an AI platform for building, training, and serving specialized models across text, image, embedding, audio, and multimodal workloads.

What you will do

  • Own the roadmap, strategy, and success metrics for inference and platform products across API, UI, and CLI.
  • Work with AI-native startups and enterprises running production inference and convert recurring customer pain into product capabilities.
  • Manage account, onboarding, quota, rate-limit, billing, and fraud- and abuse-prevention surfaces.
  • Partner with inference, infrastructure, product marketing, sales, and field teams to launch platform capabilities, pricing, documentation, cookbooks, and enablement.
  • Define and evaluate tradeoffs involving rate limits, latency, throughput, reliability, cost, and GPU utilization.

Requirements

  • 2–8+ years of product management experience building technical or developer-facing products.
  • Strong technical background through a CS/EE degree, production engineering experience, or equivalent expertise.
  • Familiarity with the inference lifecycle, including model serving, latency, throughput, and production cost tradeoffs.
  • End-to-end product ownership from strategy and specifications through launch and metrics.
  • Excellent written communication and comfort working through ambiguity with a bias toward shipping and learning.
  • Hands-on experience building on an LLM API and understanding latency, cost, and reliability tradeoffs.

Nice to have

  • Experience with self-serve or product-led growth products, usage-based billing, quotas, and rate limits.
  • Experience with trust and safety, fraud, or abuse prevention at scale.
  • Understanding of GPU economics, utilization, batching, latency SLAs, and margin.
  • Experience with cloud or developer platforms, metered pricing, identity, accounts, or organization management.
  • Early-stage startup or founding experience.

Culture & Benefits

  • Work on AI infrastructure and low-latency inference at production scale.
  • Collaborate with engineers and AI researchers on emerging technology.
  • High ownership and direct product impact in a fast-growing environment.
  • Compensation includes equity.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →