Назад
Company hidden
2 дня назад

Software Engineer, Machine Learning Platform - Gen AI

130 600 - 192 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Software Engineer, Machine Learning Platform - Gen AI (Python/LLM/GPU): Building production infrastructure for open-weight model serving, batch inference, and fine-tuning with an accent on GPU performance, autoscaling, observability, and cost efficiency. Focus on designing distributed systems for real-time inference, high-throughput pipelines, SFT/DPO/LoRA fine-tuning, and reliable GenAI platform primitives.

Location: San Francisco, CA; Sunnyvale, CA; or Seattle, WA, United States

Base salary: $130,600–$192,000 USD annually, plus potential equity grants.

Company

hirify.global operates a GenAI platform supporting AI-powered products, agents, automation, and personalization across hirify.global, Wolt, and Deliveroo.

What you will do

  • Build production infrastructure that moves GenAI applications from prototypes to reliable products.
  • Develop open-weight model serving, including real-time GPU endpoints, batch inference, and fine-tuning with SFT, DPO, and LoRA.
  • Design scalable systems for model serving, GPU autoscaling, batch pipelines, backend services, and observability.
  • Optimize GPU inference cost, latency, throughput, and utilization while maintaining reliability, fallback controls, and SLOs.
  • Build platform capabilities covering gateways, evaluations, guardrails, cost attribution, monitoring, and operational playbooks.
  • Partner with ML engineers, product engineers, data scientists, and platform teams across hirify.global, Wolt, and Deliveroo.

Requirements

  • Bachelor’s, master’s, or PhD in Computer Science or equivalent experience.
  • 3+ years of industry software engineering experience.
  • Strong backend engineering fundamentals, especially Python and distributed systems.
  • Experience building and operating production services, APIs, data pipelines, or ML infrastructure at scale.
  • Hands-on production experience with LLM inference or fine-tuning of open-weight models, including serving, batching, autoscaling, GPU utilization, SFT, DPO, or LoRA.
  • Experience with observability, debugging, reliability, incident response, and performance or cost optimization; proficiency with AI coding tools across the software development lifecycle.

Nice to have

  • Experience with vLLM, SGLang, TensorRT-LLM, distributed training, GPU optimization, quantization, or high-throughput batch systems.
  • Experience with Kubernetes, AWS, GCP, GPUs, serverless or elastic GPU platforms, and internal developer platforms.
  • Experience with LLM gateways, model routing, cost attribution, AI agents, MCP servers, evaluation systems, observability, tracing, RAG, search, or vector databases.

Culture & Benefits

  • Regular employee benefits include a 401(k) plan with employer matching, medical, dental, and vision insurance.
  • Benefits include 16 weeks of paid parental leave, paid time off, paid sick leave, 11 paid holidays, disability and life insurance, and family-forming assistance.
  • Additional benefits include wellness, commuter, and mental health programs.
  • Salaried employees receive flexible paid time off and 80 hours of paid sick time per year.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →