Назад
Company hidden
1 день назад

Senior Staff Engineer (AI/ML)

240 000 - 265 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Staff Engineer (AI/ML): Building trustworthy AI evaluation systems and data pipelines for a regulated healthcare platform with an accent on LLM judges, statistical regression gates, dataset versioning, and red-team testing. Focus on designing scalable evaluation methodology, solving ambiguous cross-team infrastructure problems, and delivering production systems that determine whether AI models are safe and effective for patients.

Location: US Remote

Salary: $240,000–$265,000 annual base salary for US-based candidates, plus equity; ranges may vary by location outside the US.

Company

Hims & Hers is a public health and wellness platform delivering personalized healthcare from diagnosis through treatment and delivery.

What you will do

  • Set technical direction for AI evaluation systems, including metrics, LLM judges, scorers, statistical regression gates, and failure tracking.
  • Design and scale ingestion, transformation, dataset versioning, labeling, calibration, feature, and benchmark pipelines.
  • Define evaluation approaches for new AI services and replace inconsistent manual reviews with automated statistical gates.
  • Build adversarial and red-team test suites and failure taxonomies to identify safety and edge-case issues before deployment to patients.
  • Lead multi-team technical initiatives, resolve disagreements, and modernize core evaluation and data systems.
  • Partner with ML engineering, data science, platform engineering, clinical, legal, and product teams while mentoring engineers.

Requirements

  • 10+ years of experience in ML infrastructure, data engineering, or evaluation and testing systems.
  • Hands-on experience designing and calibrating LLM judges and scorers, statistically sound regression testing, human-label agreement measurement, and adversarial evaluation.
  • Strong experience with dataset versioning, feature and benchmark pipelines, labeling and calibration workflows, and high-throughput data systems.
  • Experience leading multi-team projects, resolving technical disagreements, and establishing reusable standards.
  • Strong Python skills and statistical fluency sufficient to design and defend production testing frameworks.
  • Must be based in the United States for the stated salary range; location-specific compensation may apply elsewhere.

Nice to have

  • Experience with Databricks, MLflow, Unity Catalog, or similar data and evaluation platforms.
  • Experience building reporting tools for non-engineering users.
  • Experience in a regulated industry such as healthcare, fintech, or life sciences.
  • Company-wide talks or technical write-ups that influenced other teams.

Culture & Benefits

  • Competitive salary and equity compensation.
  • Unlimited PTO, company holidays, and quarterly mental health days.
  • Medical, dental, and vision coverage, plus parental leave.
  • Employee Stock Purchase Program and 401(k) employer matching.
  • Offsite team retreats and a flexible remote work approach.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →