Назад
Company hidden
5 часов назад

Forward Deployed Engineer (AI)

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Forward Deployed Engineer (AI): Deploying and integrating model APIs, GPU cloud infrastructure, and agent sandbox environments into customer AI applications with an accent on production inference, distributed systems, and customer-specific solutions. Focus on debugging APIs, networking, containers, GPUs, and production workloads while translating customer requirements into reusable tooling and platform improvements.

Location: San Mateo, United States; on-site

Company

hirify.global provides a unified AI and agent cloud platform with model APIs, GPU cloud infrastructure, and agent sandbox environments for building, deploying, and scaling production AI systems.

What you will do

  • Work directly with customers to understand AI products, technical requirements, and business goals.
  • Deploy and integrate model APIs, inference endpoints, dedicated model hosting, GPU cloud infrastructure, and agent sandbox environments.
  • Debug production issues across APIs, networking, containers, GPUs, and distributed systems.
  • Build customer-specific solutions, integrations, demos, prototypes, reusable tooling, automation, and reference implementations.
  • Support proof-of-concepts and help customers move AI workloads into production.
  • Translate customer feedback into product requirements and collaborate with internal engineering teams on platform improvements.

Requirements

  • Strong software engineering fundamentals.
  • Proficiency in Python and at least one additional programming language.
  • Experience building or deploying cloud-native applications.
  • Familiarity with Docker, Kubernetes, Linux, and networking concepts.
  • Strong production debugging, problem-solving, communication, and customer collaboration skills.
  • Ability to work with ambiguity and move quickly in a startup environment.

Nice to have

  • Experience with AI/ML infrastructure, LLMs, inference systems, vLLM, SGLang, Ray, or distributed serving frameworks.
  • Experience with GPUs, CUDA, large-scale compute systems, AI agents, coding agents, or sandboxed execution environments.
  • Background in solutions engineering, developer relations, or customer-facing engineering.
  • Prior startup or early-engineer experience.
  • Professional working proficiency in Mandarin Chinese.

Culture & Benefits

  • Work on AI infrastructure used by AI companies and developers.
  • Solve technical problems across inference, GPUs, agent systems, and distributed computing.
  • Direct influence on customer outcomes and product direction.
  • Travel to customer sites when necessary.
  • Competitive pay package, employer-covered premium medical, dental, and vision insurance, 401(k) plan, and free office meals.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →