Назад
Company hidden
6 дней назад

Staff ML Infrastructure Engineer (MLOps)

218 000 - 285 000$
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Staff ML Infrastructure Engineer (MLOps): Architecting and operating production-grade ML systems at scale with an accent on distributed training pipelines, feature stores, and high-throughput inference serving. Focus on optimizing GPU utilization, designing zero-downtime deployment systems, and industrializing AI for a high-growth retail platform.

Location: Palo Alto, California, United States

Salary: $218,000–$285,000 (base) + bonus and stock

Company

hirify.global is a tech-driven retail company disrupting the industry by leveraging AI, analytics, and automation to provide high-quality essentials through a direct-to-consumer model.

What you will do

  • Own the end-to-end technical design of the ML platform, including model training, serving, feature pipelines, and monitoring.
  • Implement a "paved road" developer experience to enable Data Scientists and AI Researchers to move from idea to production with minimal friction.
  • Establish and uphold engineering standards for CI/CD in ML, Infrastructure as Code (IaC), and model versioning.
  • Lead the evaluation and selection of core platform components, including inference runtimes and feature stores.
  • Optimize GPU utilization, model batching, and cloud cost controls to maximize performance per dollar.
  • Architect scalable serving infrastructure capable of handling traffic surges and seasonal spikes with automated recovery.

Requirements

  • 8+ years of industry experience, with at least 4+ years focused on ML Infrastructure or MLOps.
  • Deep expertise in AWS, Kubernetes (EKS), Docker, and Terraform/Pulumi.
  • Hands-on mastery of ML frameworks such as PyTorch, TensorFlow, Kubeflow, or SageMaker.
  • Experience building Feature Stores and high-throughput data pipelines using Spark, Flink, or Kafka.
  • Expert knowledge of CI/CD for ML, including blue-green and canary deployment strategies.
  • Must be based in Palo Alto, California, United States

Culture & Benefits

  • Opportunity to work in a mission-driven team reshaping the retail industry with a focus on sustainability.
  • High-growth startup environment for builders and innovators who thrive in ambiguity.
  • Competitive compensation package including base salary, bonuses, and stock options.
  • Culture of rigorous operational excellence and technical mentorship.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →