2 дня назад
Principal AI Evaluations Platform Engineer
183 800 - 309 700CHF
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal AI Evaluations Platform Engineer (AI/Python): Building and operating reliable evaluation infrastructure for large-scale model training, including benchmark configurations, graders, evaluation runners, monitoring, and scheduling with an accent on distributed systems reliability and operational excellence. Focus on debugging cross-stack failures, improving evaluation throughput, and supporting LLM evaluation workloads during European business hours.
Location: London, United Kingdom, or Zurich, Switzerland. Candidates must be local to the applicable office and work in the office four days per week. The role includes on-call coverage during European business hours.
Salary: CHF 183,800–309,700 per year in Switzerland, with possible benefits and additional compensation.
Company
AI builds artificial intelligence systems across agents, applications, services, and infrastructure.
What you will do
- Build, operate, and extend the evaluation platform supporting large-scale model training.
- Develop benchmark configurations, problem sets, graders, and evaluation runners.
- Improve evaluation throughput, scheduling efficiency, monitoring, alerting, and dashboards.
- Partner with researchers and training teams across European offices to define requirements, coordinate launches, and unblock experiments.
- Contribute to technical design, code reviews, operational documentation, and on-call runbooks.
Requirements
- Bachelor’s degree and 4+ years of technical engineering experience, or a master’s degree and 2+ years of experience, or equivalent experience.
- Professional coding experience with Python, C++, Go, or Rust; strong Python proficiency and experience with a large shared typed codebase.
- Experience operating production distributed systems and participating in a formal on-call rotation.
- Ability to debug distributed-system failures involving scheduling, resource exhaustion, networking, and process-level faults.
- Availability for scheduled on-call coverage during European business hours and ability to work from the London or Zurich office four days per week.
Nice to have
- Experience with GPU workloads, distributed training, or large-scale inference infrastructure.
- Experience with Ray, Kubernetes, Slurm, or comparable distributed compute frameworks and schedulers.
- Experience with Datadog, Prometheus, Grafana, or similar observability and alerting tools.
- Familiarity with LLM evaluation, benchmarking, or reinforcement learning infrastructure.
- Experience reducing operational burden through reliability improvements, automation, or tooling.
Culture & Benefits
- Collaborative work with researchers and training teams across European offices.
- Focus on inclusion, respect, integrity, accountability, and a growth mindset.
- Benefits and additional compensation may be available.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Principal AI Engineer (AI)
104 800 - 165 700€
5 дней назад
Principal AI Engineer (Agents and Context)
154 000 - 243 600CAD
Anthropic
5 дней назад
Applied AI Engineer (AI)
225 000 - 255 000GBP
4 дня назад
Senior Software Engineer (Forward Deployed) (AI)
110 000 - 162 000GBP
Anthropic
5 дней назад
Applied AI Architect, Startups
165 000 - 190 000GBP
3 дня назад