Назад
Company hidden
1 час назад

Senior Machine Learning Infrastructure Engineer (ML Infrastructure)

160 000 - 200 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US/SK/Sweden +2 еще
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Machine Learning Infrastructure Engineer (ML Infrastructure): Building scalable systems, data pipelines, and GPU infrastructure for training, inference, deployment, and monitoring of machine learning models at scale with an accent on distributed systems, model versioning, experiment tracking, and cloud-native orchestration. Focus on managing petabyte-scale data, optimizing distributed ML workloads, improving CI/CD and platform reliability, and developing monitoring, logging, and alerting systems.

Location: Santa Clara, California, United States; hybrid workplace

Salary: $160,000–$200,000 per year, hirify.global cash and equity compensation

Company

hirify.global is a Physical AI company developing AI-based virtual driver software for factory-built autonomous trucks, with operations in the United States and Europe.

What you will do

  • Design and develop scalable, high-performance systems for ML model training, inference, deployment, and monitoring.
  • Build data pipelines, model versioning systems, experiment tracking frameworks, and large-scale ML datasets.
  • Manage distributed systems, storage solutions, and GPU clusters optimized for machine learning workloads.
  • Improve CI/CD workflows and platform usability in collaboration with ML researchers and engineers.
  • Implement monitoring, logging, alerting, and reliability practices for high-availability ML infrastructure.
  • Mentor junior engineers and drive technical and quality improvements under the company’s QMS.

Requirements

  • PhD or MS in Computer Science, Electrical Engineering, or a related field; PhD new graduates or master’s graduates with 3+ years of software engineering experience are considered.
  • Experience focused on ML infrastructure or distributed systems, with strong software design and development skills.
  • Proficiency in Python, C++, and SQL.
  • Deep understanding of containerization, orchestration, distributed ML workloads, and experiment tracking tools including Docker, Kubernetes, Kubeflow, and MLflow.
  • Experience deploying and managing resources across AWS, GCP, or on-premises environments.
  • Proficiency with PyTorch or another deep learning framework, hirify.global data pipeline tools such as Apache Airflow or Prefect; strong knowledge of databases and storage solutions.

Nice to have

  • Familiarity with convolutional neural networks and Transformer models.
  • Experience with MLOps pipelines, Ray, or other distributed computing frameworks.
  • Experience working with autonomous vehicles or robotics.

Culture & Benefits

  • Work in a future-oriented and innovative autonomous vehicle field.
  • Opportunities for personal and professional development.
  • Catered free lunch, unlimited snacks, and beverages.
  • Competitive salary and benefits package, including a 401(k) plan.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →