7 дней назад
DevOps / SRE Engineer (Sheffield)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
DevOps / SRE Engineer (Kubernetes/AWS/Azure): Building and operating secure, scalable cloud infrastructure and Kubernetes platforms for a digital payments and insights platform with an accent on infrastructure automation, observability, and platform reliability. Focus on incident response, defining SLIs/SLOs, improving CI/CD, and reducing operational toil through automation.
Location: Hybrid role based in the Sheffield office, at least 3 days per week. Candidates must already have the right to work in the UK.
Company
is a multi-award-winning digital payments and insights platform for the automotive industry, supporting more than 5,000 automotive retailers across the UK and Europe.
What you will do
- Design, build, and maintain secure, scalable cloud infrastructure across AWS and Azure.
- Manage and enhance the Kubernetes EKS platform for reliable modern applications.
- Develop and maintain Infrastructure as Code with Terraform and Helm.
- Improve CI/CD pipelines using Argo Workflows, ArgoCD, and GitHub Actions.
- Lead and participate in incident response, on-call activities, major incident coordination, and blameless post-incident reviews.
- Drive monitoring, alerting, observability, SLIs/SLOs, automation, and developer experience improvements.
Requirements
- Proven experience in DevOps, SRE, or Platform Engineering.
- Strong hands-on experience running Kubernetes in production.
- Experience with AWS and/or Azure, Terraform, and Infrastructure as Code automation.
- Experience with production incidents, on-call operations, monitoring, alerting, and observability.
- Ability to diagnose complex distributed systems issues and communicate calmly under pressure.
- English is the shared working language, and candidates must already have the right to work in the UK; visa sponsorship is not available.
Nice to have
- Experience with Grafana Cloud, IRM, or other incident management tools.
- Familiarity with Argo Workflows, ArgoCD, and defining SLIs/SLOs at scale.
- Platform Engineering experience or experience improving incident management processes.
Culture & Benefits
- International team with a strong focus on reliability, continuous improvement, automation, and reducing toil.
- Competitive salary, company bonus scheme, private healthcare, and Medicash.
- 26 days of holiday plus bank holidays and volunteer days.
- Salary sacrifice pension with Aviva, Cycle to Work, electric vehicle, and nursery fee schemes.
- Annual wellbeing and development allowances of £250 and £150 respectively.
- Paid parental leave of up to four months for primary carers and one month for secondary carers.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Site Reliability Engineer (AI Platform)
69 768 - 96 900€
13 дней назад
Site Reliability Engineer (Cloud Banking)
9 дней назад
Site Reliability Engineer (DataCosmos)
7 дней назад
Associate Site Reliability Engineer
10 дней назад
Site Reliability Engineer (AWS/Kubernetes)
90 000 - 120 000GBP
12 дней назад