Назад
9 часов назад

Platform Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Platform Engineer (Kubernetes/AWS): Building and operating the compute layer underpinning DeepL's AI products and research with an accent on hybrid infrastructure (on-prem and cloud) and GPU clusters. Focus on designing production-grade Kubernetes clusters, unifying workloads across AWS and on-prem hardware, and improving platform reliability.

Location: Hybrid in London (office attendance twice a week)

Company

DeepL is a global AI product and research company focused on building secure, intelligent solutions to complex business problems through its Language AI platform.

What you will do

  • Own the reliability, capacity, and cost efficiency of CPU/GPU compute infrastructure across AWS and on-prem.
  • Shape the platform's architecture and drive technical decisions for optimal system outcomes.
  • Design, build, and operate production-grade Kubernetes clusters through their full lifecycle.
  • Unify workload execution across on-prem hardware and AWS to support product and research teams.
  • Define IaC and platform standards using Terraform and GitOps tooling like ArgoCD.
  • Mentor engineers and lead incident response to ensure long-term platform resilience.

Requirements

  • Deep, hands-on expertise in designing and operating Kubernetes clusters at scale.
  • Strong experience with AWS or on-prem infrastructure (bare-metal or data center).
  • Proficiency in Linux and networking for deep-level debugging from container to network edge.
  • Experience with Infrastructure as Code (Terraform) and GitOps (ArgoCD) in production.
  • Software engineering skills in Go or Python for building internal tools.
  • Proven track record of technical ownership and reliability engineering discipline.

Nice to have

  • Experience with GPU infrastructure for training or inference at scale.
  • Knowledge of distributed storage systems, such as Ceph.
  • Experience providing platform services to researchers or data scientists.
  • Networking expertise at the BGP level in hybrid or on-prem environments.

Culture & Benefits

  • Hybrid work schedule with office attendance twice a week and flexible hours.
  • Virtual Shares providing employees with a stake in the company's growth.
  • Monthly "Hack Fridays" for pursuing passionate projects and cross-team collaboration.
  • 30 days of annual leave plus public holidays and access to mental health resources.
  • Diverse, international environment with a team representing over 90 nationalities.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →