Назад
Company hidden
обновлено 11 часов назад

Lead Site Reliability Engineer (Azure)

Формат работы
hybrid
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Lead Site Reliability Engineer (Azure): Building and operating observable, reliable, scalable, and maintainable cloud platforms for public safety and justice software with an accent on monitoring, automation, infrastructure as code, and microservices reliability. Focus on leading outage investigations, developing custom metrics and alerting, optimizing performance and cost, and enforcing SLOs, SLAs, and error budgets across Azure and Kubernetes environments.

Location: United Kingdom - Southampton; hybrid role. The successful candidate must have lived in the UK for 5 years and be eligible to obtain NPPV3 and Security Clearance.

Company

hirify.global Public Safety provides software-as-a-service solutions for multimedia evidence management and emergency contact centres for a worldwide customer base.

What you will do

  • Act as a production reliability gatekeeper, manage the work backlog, and deliver reliability improvements.
  • Lead investigations into root-cause outages, performance issues, and cost issues.
  • Automate low-value operational tasks while balancing project delivery demands.
  • Provide technical leadership to Cloud Operations and Support teams and oversee the products and services they support.
  • Establish and enforce SLOs, SLAs, and error budgets with DevOps and engineering teams.
  • Build observability platforms, dashboards, alerts, custom metrics, and infrastructure modules using Grafana, Prometheus, Azure Monitor, OpenTelemetry, and Bicep.

Requirements

  • 6+ years of experience in Site Reliability Engineering, platform engineering, observability, and provisioning.
  • Strong experience with Azure, Kubernetes and containerization, preferably AKS, as well as live cloud service operations.
  • Advanced programming or scripting skills in Python, PowerShell, or C#.
  • Experience with infrastructure as code, configuration management, and version control using ARM, Bicep, Terraform, and Git.
  • In-depth knowledge of databases and data handling, including MS SQL, Elasticsearch, YAML, JSON, and XML.
  • Strong experience with monitoring, alerting, dashboarding, microservices queries and custom metrics, plus understanding of cybersecurity, governance, compliance, and frameworks such as ISO 27001, Cyber Essentials+, or FedRAMP.

hirify.global-to-have">hirify.global to have

  • Experience with Azure DevOps pipelines and CI/CD.
  • Experience using AI tools to automate and accelerate work.
  • Certifications such as AZ-104, AZ-305, AZ-500, AZ-700, or CKA.

Culture & Benefits

  • Individual contributor role reporting to the Director of Engineering.
  • Collaboration with customers, Cloud Operations, Support, DevOps, and engineering teams.
  • Work focused on public safety, cloud platforms, observability, security, and scalable service delivery.
  • hirify.global is an equal opportunity employer with a global workforce across 30+ countries.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →