Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Platform Engineer (Kubernetes/AWS): Building and operating the compute layer underpinning DeepL's AI products and research with an accent on hybrid infrastructure (on-prem and cloud) and GPU clusters. Focus on designing production-grade Kubernetes clusters, unifying workloads across AWS and on-prem hardware, and improving platform reliability.
Location: Hybrid in London (office attendance twice a week)
Company
DeepL is a global AI product and research company focused on building secure, intelligent solutions to complex business problems through its Language AI platform.
What you will do
- Own the reliability, capacity, and cost efficiency of CPU/GPU compute infrastructure across AWS and on-prem.
- Shape the platform's architecture and drive technical decisions for optimal system outcomes.
- Design, build, and operate production-grade Kubernetes clusters through their full lifecycle.
- Unify workload execution across on-prem hardware and AWS to support product and research teams.
- Define IaC and platform standards using Terraform and GitOps tooling like ArgoCD.
- Mentor engineers and lead incident response to ensure long-term platform resilience.
Requirements
- Deep, hands-on expertise in designing and operating Kubernetes clusters at scale.
- Strong experience with AWS or on-prem infrastructure (bare-metal or data center).
- Proficiency in Linux and networking for deep-level debugging from container to network edge.
- Experience with Infrastructure as Code (Terraform) and GitOps (ArgoCD) in production.
- Software engineering skills in Go or Python for building internal tools.
- Proven track record of technical ownership and reliability engineering discipline.
Nice to have
- Experience with GPU infrastructure for training or inference at scale.
- Knowledge of distributed storage systems, such as Ceph.
- Experience providing platform services to researchers or data scientists.
- Networking expertise at the BGP level in hybrid or on-prem environments.
Culture & Benefits
- Hybrid work schedule with office attendance twice a week and flexible hours.
- Virtual Shares providing employees with a stake in the company's growth.
- Monthly "Hack Fridays" for pursuing passionate projects and cross-team collaboration.
- 30 days of annual leave plus public holidays and access to mental health resources.
- Diverse, international environment with a team representing over 90 nationalities.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →