3 дня назад
DevOps / SRE Engineer (London)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
DevOps / SRE Engineer (London) (Cloud Infrastructure/Kubernetes): Building and operating secure, scalable cloud infrastructure and Kubernetes platforms for a digital payments and insights platform with an accent on reliability, observability, and infrastructure automation. Focus on managing EKS, developing Terraform and Helm infrastructure, improving CI/CD, coordinating incident response, and defining SLIs/SLOs.
Location: Hybrid, based in the London office at least 3 days a week; candidates must already have the right to work in the UK. Visa sponsorship is not available.
Company
is a digital payments and insights platform for the automotive industry, helping automotive retailers offer flexible payment options for motoring bills, accessories, and servicing.
What you will do
- Design, build, and maintain secure, scalable cloud infrastructure across AWS and Azure.
- Manage and enhance the Kubernetes EKS platform for reliable modern applications.
- Develop and maintain Infrastructure as Code using Terraform and Helm.
- Improve CI/CD pipelines with Argo Workflows, ArgoCD, and GitHub Actions.
- Lead and participate in incident response, on-call activities, major incident coordination, and blameless post-incident reviews.
- Drive monitoring, alerting, observability, SLI/SLO implementation, automation, and developer experience improvements.
Requirements
- Proven experience in DevOps, SRE, or Platform Engineering.
- Strong hands-on experience running Kubernetes in production.
- Experience with AWS and/or Azure, Terraform, and Infrastructure as Code automation.
- Experience with production incidents, on-call responsibilities, monitoring, alerting, and observability.
- Ability to diagnose and resolve complex distributed-systems issues, with a strong focus on automation and reducing toil.
- English is the shared working language, and candidates must already have the right to work in the UK.
Nice to have
- Experience with Grafana Cloud, IRM, or other incident management tools.
- Familiarity with Argo Workflows or ArgoCD.
- Experience defining SLIs/SLOs at scale and improving incident management processes.
- Background in Platform Engineering.
Culture & Benefits
- International engineering environment with collaboration across the UK and Europe.
- Company bonus scheme and private healthcare with Medicash.
- 26 days of holiday plus bank holidays and volunteer days.
- Salary sacrifice pension with Aviva and cycle-to-work, electric vehicle, and nursery fee schemes.
- Annual wellbeing and development allowances of £250 and £150 respectively.
- Paid parental leave of up to four months for primary carers and one month for secondary carers.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Senior Site Reliability Engineer (AI Platform)
69 768 - 96 900€
6 дней назад
Site Reliability Engineer (DataCosmos)
4 дня назад
NOC Engineer / SRE
9 дней назад
Senior Site Reliability Engineer, IaaS (Kubernetes)
7 дней назад
Site Reliability Engineer (AWS/Kubernetes)
90 000 - 120 000GBP
9 дней назад