Назад
Company hidden
2 дня назад

Principal AI Evaluations Platform Engineer

183 800 - 309 700CHF
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK/Switzerland
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal AI Evaluations Platform Engineer (AI/Python): Building and operating reliable evaluation infrastructure for large-scale model training, including benchmark configurations, graders, evaluation runners, monitoring, and scheduling with an accent on distributed systems reliability and operational excellence. Focus on debugging cross-stack failures, improving evaluation throughput, and supporting LLM evaluation workloads during European business hours.

Location: London, United Kingdom, or Zurich, Switzerland. Candidates must be local to the applicable office and work in the office four days per week. The role includes on-call coverage during European business hours.

Salary: CHF 183,800–309,700 per year in Switzerland, with possible benefits and additional compensation.

Company

hirify.global AI builds artificial intelligence systems across agents, applications, services, and infrastructure.

What you will do

  • Build, operate, and extend the evaluation platform supporting large-scale model training.
  • Develop benchmark configurations, problem sets, graders, and evaluation runners.
  • Improve evaluation throughput, scheduling efficiency, monitoring, alerting, and dashboards.
  • Partner with researchers and training teams across European offices to define requirements, coordinate launches, and unblock experiments.
  • Contribute to technical design, code reviews, operational documentation, and on-call runbooks.

Requirements

  • Bachelor’s degree and 4+ years of technical engineering experience, or a master’s degree and 2+ years of experience, or equivalent experience.
  • Professional coding experience with Python, C++, Go, or Rust; strong Python proficiency and experience with a large shared typed codebase.
  • Experience operating production distributed systems and participating in a formal on-call rotation.
  • Ability to debug distributed-system failures involving scheduling, resource exhaustion, networking, and process-level faults.
  • Availability for scheduled on-call coverage during European business hours and ability to work from the London or Zurich office four days per week.

Nice to have

  • Experience with GPU workloads, distributed training, or large-scale inference infrastructure.
  • Experience with Ray, Kubernetes, Slurm, or comparable distributed compute frameworks and schedulers.
  • Experience with Datadog, Prometheus, Grafana, or similar observability and alerting tools.
  • Familiarity with LLM evaluation, benchmarking, or reinforcement learning infrastructure.
  • Experience reducing operational burden through reliability improvements, automation, or tooling.

Culture & Benefits

  • Collaborative work with researchers and training teams across European offices.
  • Focus on inclusion, respect, integrity, accountability, and a growth mindset.
  • Benefits and additional compensation may be available.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →