9 дней назад
SRE (AWS/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
SRE (AWS/Kubernetes): Owning Aqemia’s cloud platform across infrastructure-as-code, GitOps delivery, observability, security, and FinOps for large-scale scientific computation with an accent on GPU fleet planning, reliability, and cost control. Focus on designing fully automated AWS infrastructure, operating Kubernetes through GitOps, and building secure CI/CD and internal tooling for engineering, ML, and scientific teams.
Location: Paris, hybrid, with up to two remote days per week. also has an office in London.
Company
is a drug invention company developing new medicines through its QEMI platform, which combines physics-based modeling, statistical mechanics, and generative AI.
What you will do
- Own AWS infrastructure end to end using OpenTofu and Terragrunt, with all changes managed as code.
- Operate and evolve Kubernetes workloads through GitOps with ArgoCD, Helm, and Kustomize.
- Manage observability, incident response, blameless postmortems, and reliability improvements.
- Track cloud costs, maintain cost allocation, and plan capacity for large GPU computing campaigns.
- Enforce platform security through patching, vulnerability follow-up, and least-privileged access management.
- Build internal tooling and CI/CD pipelines for engineering, ML, and scientific teams, while contributing to platform architecture and strategy.
Requirements
- 2–3+ years of post-degree experience in platform or infrastructure engineering.
- Deep hands-on experience with AWS and infrastructure-as-code, including Terraform or OpenTofu and Terragrunt.
- Production experience with Kubernetes and GitOps delivery using ArgoCD, Helm, and Kustomize.
- Experience building and maintaining CI/CD pipelines with GitHub Actions or GitLab.
- Experience with cloud security practices and least-privileged access management.
Nice to have
- MLOps experience with training or inference pipelines, model lifecycle management, or workflow orchestrators.
- GPU capacity planning, autoscaling, spot strategies, or quota management.
- Experience with AI-driven or data-intensive workflows.
- Experience with a cloud provider beyond AWS, such as GCP.
Culture & Benefits
- Work at the intersection of AI, physics-based modeling, biology, medicinal chemistry, and drug discovery.
- Contribute to a growing pipeline of proprietary therapeutic programs and pharmaceutical partnerships.
- Flexible hybrid work with up to two remote days per week.
- Join a small engineering team with end-to-end ownership of the platform.
- Work for a company backed by $100M from European and international investors.
Hiring process
- Initial discussion with Talent Acquisition.
- Hiring Manager interview followed by a technical deep-dive with the team.
- Cultural fit interview with the co-founder and COO, followed by a final interview with the co-founder and CEO.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
DevOps / SRE Engineer (London)
11 дней назад
DevOps / SRE Engineer (Sheffield)
ClickHouse
14 дней назад
Infrastructure Engineer (AWS)
90 000 - 160 000€
14 дней назад
Senior Site Reliability Engineer (AI Platform)
69 768 - 96 900€
14 дней назад
Senior Site Reliability Engineer, PaaS (Kubernetes)
69 768 - 96 900€
13 дней назад