Назад
Company hidden
обновлено 2 часа назад

Senior Site Reliability Engineer (Kubernetes)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
CR
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (Kubernetes/AWS/Azure): Building, operating, and improving highly available production infrastructure and internal platforms with an accent on Kubernetes, cloud environments, Terraform, databases, and observability. Focus on automating operational work, troubleshooting distributed-system failures, responding to incidents, and strengthening resilience through upgrades, capacity planning, and disaster recovery.

Location: Brno, Czech Republic; hybrid 3+2 schedule with at least three days per week in the office, including mandatory Tuesdays and Wednesdays

Company

hirify.global develops simple, powerful, and secure solutions that help customers accelerate business transformation.

What you will do

  • Operate, upgrade, troubleshoot, and improve production Kubernetes clusters and workloads across AWS and Azure.
  • Manage Kubernetes platforms and components including Helm, Kustomize, operators, Istio, autoscaling, and cluster lifecycle operations.
  • Operate production database platforms including ClickHouse, Aurora, and other distributed data systems.
  • Build infrastructure and automation with Terraform, Python, Go, Bash, or similar technologies.
  • Improve observability, monitoring, logging, alerting, incident response, resilience, disaster recovery, and infrastructure lifecycle management.
  • Participate in scheduled on-call rotations, lead or support incident resolution and root-cause analysis, and collaborate with software engineering and platform teams.

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Platform Engineering, or a related field.
  • Strong hands-on experience operating, upgrading, and troubleshooting Kubernetes clusters and workloads at both application and infrastructure levels.
  • Strong hands-on experience with AWS and Azure, Linux systems administration, Terraform, and Infrastructure as Code.
  • Experience supporting customer-facing, highly available production systems and participating in high-severity incident response.
  • Strong scripting and automation skills using Python, Go, Bash, or similar languages.
  • Strong understanding of infrastructure, networking, storage, DNS, load balancing, security, distributed-system troubleshooting, and high-availability patterns.

Culture & Benefits

  • People-first environment focused on collaboration, accountability, empathy, ownership, and continuous improvement.
  • 25 vacation days, three sick days, 10 study days, and two volunteering days per year.
  • Additional holiday entitlement after five years and sabbatical leave.
  • Personal education budget of up to 48,300 CZK per year and unlimited LinkedIn Learning access.
  • Pension or life insurance contribution matching up to 3% of salary or 4,000 CZK per month, plus a meal allowance.
  • English and Czech classes, Multisport card, employee assistance, referral and appreciation programs, and donation matching.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →