Назад
Company hidden
обновлено 2 часа назад

Engineering Manager, Provider Ecosystem (AI)

245 000 - 295 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Engineering Manager, Provider Ecosystem (AI) (LLM infrastructure): Leading the provider ecosystem team responsible for model and endpoint integrations, routing, and inference quality across hundreds of providers and millions of requests with an accent on distributed systems, API design, and multimodal inference. Focus on growing the engineering team, launching new providers, building benchmark-driven routing decisions, and solving reliability and performance challenges across heterogeneous endpoints.

Location: Remote (US)

Salary: $245K–$295K annually plus equity

Company

hirify.global provides an AI routing and infrastructure layer that enables enterprises to access, manage, and optimize large language models across multiple providers.

What you will do

  • Lead and grow the Provider Ecosystem engineering team, set the roadmap, conduct code reviews, and participate in incident response.
  • Own provider adapters, endpoint lifecycles, and routing across hundreds of providers and thousands of endpoints.
  • Act as the technical counterpart for model labs and inference providers, coordinating integrations and rapid day-zero launches.
  • Expand platform capabilities beyond text into image, video, audio, embeddings, realtime inference, and batch inference.
  • Own evaluations, benchmarks, and production telemetry used to guide routing decisions and improve provider quality.

Requirements

  • 4+ years of engineering management experience, including experience building or scaling a team during rapid growth.
  • Strong foundation in distributed systems, API design, and LLM inference, including streaming, tool calling, prompt caching, quantization, throughput, and latency trade-offs.
  • Hands-on coding experience and comfort using AI-assisted development tools in a shipping-focused engineering organization.
  • Experience recruiting and retaining strong engineers, communicating clearly, and driving initiatives with a high degree of autonomy.
  • Ability to work directly with external model labs and providers on technical integrations and requirements.

Nice to have

  • Experience at an inference provider, model lab, GPU cloud, or with products integrating multiple LLM APIs.
  • Experience with TypeScript, Cloudflare Workers, Postgres, ClickHouse, GCP, or Vercel.
  • Background in evaluations, benchmarking, routing, or load-balancing systems.
  • Experience shipping multimodal APIs, realtime inference, or batch inference.
  • Experience with vendor security and compliance reviews, including NDAs, DPAs, and SOC 2.

Culture & Benefits

  • Remote work from the US.
  • Lean engineering organization focused on speed, clarity, and practical execution.
  • Opportunity to work with novel and stealth inference offerings from model labs and providers.
  • Equity included in the compensation package.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →