Назад
2 дня назад

AI Field Engineer (Enterprise)

Формат работы
remote (только USA)/hybrid
Тип работы
fulltime
Грейд
senior
Английский
c1
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

AI Field Engineer (Enterprise): Building and deploying production-grade generative AI systems for enterprise customers with an accent on inference optimization, fine-tuning pipelines, and scalable architecture. Focus on bridging the gap between complex AI research and real-world business outcomes through hands-on coding, stakeholder management, and technical discovery.

Location: Must be based in the USA. This role involves both remote work and on-site customer engagements.

Company

Fireworks AI is a high-growth, Series C startup building the industry's fastest and most scalable generative AI inference infrastructure, founded by veterans from Meta PyTorch and Google Vertex AI.

What you will do

  • Build end-to-end POCs and MVPs directly within customer codebases and infrastructure.
  • Architect inference foundations and tune deployments to meet latency, throughput, and cost targets.
  • Guide customers through model selection, fine-tuning strategies (SFT, DPO, RFT), and evaluation methodologies.
  • Lead technical discovery and maintain relationships with both ML engineers and executive stakeholders.
  • Translate customer pain points and deployment patterns into actionable product roadmap improvements.
  • Spend time on-site with customers to build trust and accelerate production deployments.

Requirements

  • 5+ years in a customer-facing technical role such as Field Engineer, Solutions Architect, or ML Engineer.
  • Strong Python proficiency with experience shipping production-grade code.
  • Hands-on experience with the LLM stack, including inference trade-offs and fine-tuning workflows.
  • Experience with cloud infrastructure (AWS, Azure, GCP) and deploying models on GPU infrastructure.
  • Familiarity with Kubernetes and infrastructure engineering.
  • Exceptional communication skills with the ability to present to VPs and debug technical issues with engineers.

Nice to have

  • 10+ years of experience in technical field or engineering roles.
  • Experience with inference serving frameworks like vLLM, SGLang, or TensorRT-LLM.
  • Track record of taking GenAI POCs to production-scale deployments.
  • Experience with hyperscaler AI platforms like Azure AI Foundry, AWS Bedrock, or GCP Vertex.
  • Experience building agentic systems or AI-native developer toolchains.

Culture & Benefits

  • Work at the forefront of AI infrastructure and low-latency inference.
  • Collaborate with world-class engineers and researchers from top AI labs.
  • High-impact environment with minimal bureaucracy and a focus on shipping results.
  • Opportunity to shape the future of how businesses harness generative AI globally.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →