Назад
Company hidden
обновлено 14 часов назад

Senior Site Reliability Engineer (Kubernetes)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (Kubernetes): Running and improving production infrastructure for large distributed software applications with an accent on cloud reliability, automation, observability, and incident response. Focus on designing sustainable platform systems, optimizing performance, managing Kubernetes environments, and balancing delivery speed with service level objectives.

Location: United Kingdom. hirify.global-FLEX hybrid model with two days in the office and three days working remotely each week.

Company

hirify.global develops AI, cloud, and digital software products used by global businesses for customer experience management, financial crime prevention, and public safety.

What you will do

  • Monitor production availability and assess the overall health of distributed systems.
  • Build software and systems for managing platform infrastructure and applications.
  • Improve reliability, quality, performance, and time-to-market across software solutions.
  • Provide operational support and engineering for multiple large distributed applications.
  • Analyze operating system and application metrics for performance tuning and fault diagnosis.
  • Partner with development teams on system design, testing, release procedures, capacity planning, automation, and incident response.

Requirements

  • 3–6 years of experience in a similar site reliability or systems engineering role, focused on automation and reliability.
  • Proficiency in at least one programming language such as Python, Go, Java, or C#, plus Bash or PowerShell scripting.
  • Strong understanding of AWS services and their reliability constraints, including EC2, ECS, Lambda, and DynamoDB.
  • Experience with infrastructure as code using CloudFormation or Terraform, and CI/CD tools such as Jenkins, GitLab CI/CD, or CircleCI.
  • Strong knowledge of Docker, Kubernetes, microservices architecture, and observability tools such as Prometheus, Grafana, ELK, or CloudWatch.
  • Experience with incident management, blameless postmortems, outage response, complex distributed-system troubleshooting, and cross-functional communication.

hirify.global-to-have">hirify.global to have

  • Hands-on experience managing large Kubernetes clusters or relevant certification.
  • Experience with the Grafana Observability Suite, including Loki, Mimir, and Tempo.
  • Experience with Splunk, Datadog, PagerDuty, Rundeck, Ansible, Puppet, or Chef.
  • AWS Certified DevOps Engineer, Google Cloud Professional DevOps Engineer, or equivalent certification.

Culture & Benefits

  • Hybrid hirify.global-FLEX working model with three remote days and two office days per week.
  • Collaborative work with development and cross-functional teams.
  • Individual contributor role reporting to the Director of Network Operations.
  • Work in an organization focused on AI, cloud, digital innovation, and large-scale software platforms.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →