Назад
Company hidden
6 дней назад

Principal Product Manager (AI Infrastructure and Orchestration)

Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal Product Manager (AI Infrastructure and Orchestration): Owning workload and agent runtime orchestration for AI deployments across Kubernetes clusters and heterogeneous accelerators with an accent on placement, lifecycle management, scaling, isolation, and platform APIs. Focus on designing multi-tenant scheduling and capacity controls, defining reliability and governance requirements, and aligning two engineering pods around a customer-facing infrastructure layer.

Location: Remote from the United States or Canada, with listed locations in Washington, Boston, Ontario, and San Francisco

Company

hirify.global provides a platform for developing, deploying, and governing predictive and generative AI for enterprise customers, including regulated, sovereign, air-gapped, and customer-managed environments.

What you will do

  • Own the deployment and workload API, including resource models, lifecycle semantics, versioning, backward compatibility, and error behavior.
  • Define workload placement, accelerator capacity, quotas, priorities, fairness, and contention behavior across tenants.
  • Shape autoscaling, cold-start and scale-to-zero policies, headroom, and cost-versus-latency controls.
  • Own agent runtime behavior, including isolation, tool-call execution, session state, restart recovery, and eviction handling.
  • Design traffic routing, request-aware load balancing, tenancy boundaries, and private connectivity to customer networks.
  • Set governance, audit, metering, packaging, SLO, error-budget, and operational-diagnostics requirements.

Requirements

  • 6+ years of product management experience in infrastructure, developer platforms, or cloud services; principal candidates typically have 9+ years.
  • At least 3 years working with Kubernetes-based or distributed-systems products and experience owning a platform layer used by other product teams.
  • Deep understanding of GPUs and accelerators, including topology-aware placement, fractional and time-sliced sharing, MIG, device plugins, and runtime plumbing.
  • Deep Kubernetes knowledge covering the API server, scheduler, controllers, CRDs, operators, admission, RBAC, resource management, and scheduling failures.
  • Experience with multi-tenancy, API products, technical writing or prototyping, and matrixed engineering collaboration without direct reports.
  • BS or MS in Computer Science or a related technical field, or equivalent hands-on software, platform, or infrastructure engineering experience.

Nice to have

  • Service networking expertise, including ingress, routing, load balancing, DNS, TLS, private connectivity, and network policy.
  • Experience with vLLM or similar serving stacks, KV caches, batching, and quantization trade-offs.
  • Experience with long-running agentic workloads, customer-managed or air-gapped deployments, regulated industries, or open-source contributions.

Culture & Benefits

  • Daily use of AI tooling, with an expectation to build prototypes, evaluation harnesses, agents, or practical tools independently.
  • Collaboration across two engineering pods and product teams building on the platform.
  • Medical, dental, and vision insurance, flexible time off, paid holidays, paid parental leave, and a global employee assistance program, depending on location and local requirements.
  • Operating principles emphasize customer impact, rigor, high standards, constructive debate, clear communication, and delivery.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →