13 дней назад
Site Reliability Engineer (AI Platform)
69 768 - 96 900€
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (AI Platform) (Kubernetes/Cloud Infrastructure/AI): Building and operating highly available production infrastructure for AI-related workloads with an accent on Kubernetes platforms, observability, CI/CD, and Infrastructure as Code. Focus on improving reliability through SLOs and capacity management, investigating production incidents, and automating complex cloud operations at scale.
Location: Paris, France; roles listed within a specific city are available only in that location, with either a hybrid-remote or in-office schedule depending on the role.
Annual base salary: €69,768–€96,900 EUR
Company
provides a search and product discovery platform that turns products, content, and business rules into actionable experiences for people, applications, and AI agents.
What you will do
- Build and operate production infrastructure for AI-related workloads and services.
- Operate and improve highly available Kubernetes-based platforms.
- Improve reliability through SLOs, observability, alerting, and capacity management.
- Investigate production incidents and implement durable fixes across networking, databases, compute, and service infrastructure.
- Improve CI/CD pipelines, deployment automation, Infrastructure as Code, and developer experience.
- Participate in on-call rotations, incident response, and operational improvements while progressively taking ownership of broader production areas.
Requirements
- Hands-on Kubernetes experience, including workloads, resource management, and production operations.
- Strong experience with Infrastructure as Code and the lifecycle of cloud infrastructure.
- Experience building and operating CI/CD pipelines and automated deployment workflows.
- Hands-on experience with at least one major cloud provider: GCP, AWS, or Azure.
- Understanding of networking, distributed systems, reliability engineering, monitoring, observability, and production troubleshooting.
- Excellent written and spoken English, with an automation mindset and ability to take ownership of production systems.
Nice to have
- Go or Python engineering experience.
- Exposure to AI/ML infrastructure and inference systems.
- Experience using coding agents, agentic workflows, and AI-assisted debugging.
Culture & Benefits
- Flexible workplace model focused on individual impact, contribution, and output.
- High-trust environment with autonomy over where and when to work, subject to the role's location and schedule.
- Options for hybrid-remote or in-office work for eligible Paris-based roles.
- Values centered on grit, trust, candor, care, and humility.
- Inclusive workplace supporting autonomy, diversity, and collaboration.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Replit
13 дней назад
Staff Site Reliability Engineer (Kubernetes/GCP)
250 000 - 325 000$
13 дней назад
Staff Service Reliability and Operational Intelligence Engineer (AI Ops)
Replit
13 дней назад
Site Reliability Engineer
210 000 - 275 000$
13 дней назад
Site Reliability Engineer (AWS/Kubernetes)
90 000 - 120 000GBP
Hippocratic AI
13 дней назад
Senior Site Reliability Engineer (AI)
13 дней назад