Назад
Company hidden
3 часа назад

API Product Engineer (AI)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
API Product Engineer (AI): Building the product platform for model APIs that lets external users ingest data, fine-tune and evaluate models, and run low-latency inference for physical robots with an accent on distributed systems, developer platforms, and reliable infrastructure. Focus on designing multi-tenant services, scaling inference from partner integrations to millions of robots, and turning rapidly evolving research capabilities into stable APIs and tools.

Location: San Francisco, United States; on-site

Company

hirify.global develops foundation models and learning algorithms that enable robots and other physically actuated devices to operate in the real world.

What you will do

  • Build Pi’s model API, including data ingestion, fine-tuning, evaluation, low-latency remote inference, partner tools, and deployment integrations.
  • Design reliable, multi-tenant infrastructure with rate limiting, isolation, backpressure, versioning, observability, and SLOs.
  • Turn partner data workflows into an intuitive product experience from upload and validation through processing, fine-tuning, and evaluation.
  • Build and operate low-latency inference systems for models controlling robots in real-world environments.
  • Work with researchers to turn new model capabilities into stable APIs, tools, documentation, and abstractions.
  • Write production-quality code and define the developer platform for general-purpose robotics.

Requirements

  • Strong software engineering fundamentals and experience building production systems.
  • Deep backend and systems experience with APIs, services, databases, caching, distributed systems, and infrastructure.
  • Experience building and scaling developer platforms, especially for model fine-tuning, inference, or compute-intensive workloads.
  • Understanding of reliability, latency, multi-tenancy, versioning, observability, and operational complexity at scale.
  • Familiarity with deploying, serving, and debugging machine learning models in production, plus strong Python skills.
  • Ability to work on-site in San Francisco.

Nice to have

  • Experience with low-latency or real-time systems, streaming, inference transport, WebSockets, QUIC, or similar technologies.
  • Experience with model serving, inference, fine-tuning, or developer-platform infrastructure.
  • Experience at an early-stage infrastructure, AI, robotics, or autonomous systems company.
  • Familiarity with Python, Postgres, ClickHouse, GCP, Kubernetes, Modal, React, and TypeScript.
  • Experience with security, authentication, authorization, or multi-tenant infrastructure.

Culture & Benefits

  • Work at the intersection of AI, infrastructure, robotics, and external partners.
  • Collaborate closely with researchers and engineers to productize rapidly evolving model capabilities.
  • Take end-to-end ownership of ambiguous platform problems and develop them into scalable systems.
  • The role focuses on product and platform development rather than customer success, partner-specific robotics, or traditional ML infrastructure.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →