1 Π΄Π΅Π½Ρ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (Kubernetes)
ΠΡΡΡ & Π‘ΠΎΠΏΡΠΎΠ²ΠΎΠ΄
ΠΠ»Ρ ΠΌΡΡΡΠ° Ρ ΡΡΠΎΠΉ Π²Π°ΠΊΠ°Π½ΡΠΈΠ΅ΠΉ Π½ΡΠΆΠ΅Π½ Plus
ΠΠΏΠΈΡΠ°Π½ΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
Π’Π΅ΠΊΡΡ:
TL;DR
Senior Site Reliability Engineer (Kubernetes/AWS/Azure): Building, operating, and improving highly available production systems and internal platforms with an accent on Kubernetes operations, cloud infrastructure, automation, and observability. Focus on troubleshooting distributed-system failures, responding to high-severity incidents, improving resilience, and managing infrastructure lifecycle across AWS and Azure.
Location: Krakow, Poland; hybrid 3+2 schedule with at least three mandatory office days per week, including Wednesdays and Thursdays.
Company
develops simple, powerful, and secure solutions that help customers accelerate business transformation.
What you will do
- Operate, maintain, upgrade, and improve production Kubernetes clusters and workloads across AWS and Azure.
- Manage Kubernetes platform components, including Helm, Kustomize, operators, Istio, autoscaling, and cluster and node lifecycle processes.
- Support production database platforms such as ClickHouse and Aurora, including performance troubleshooting and operational health.
- Build infrastructure and automation with Terraform, Python, Go, Bash, and similar technologies.
- Participate in scheduled on-call rotations, respond to production incidents, and lead or contribute to incident resolution and root-cause analysis.
- Improve observability, monitoring, logging, alerting, disaster recovery, documentation, and infrastructure lifecycle management.
Requirements
- 5+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Platform Engineering, or a related field.
- Strong hands-on production experience operating, upgrading, and troubleshooting Kubernetes clusters and workloads.
- Strong hands-on experience with AWS, Azure, Linux systems administration, Terraform, and Infrastructure as Code.
- Experience operating customer-facing, highly available production systems and participating in high-severity incident response.
- Experience with scripting and automation using Python, Go, Bash, or similar languages.
- Strong understanding of compute, networking, storage, DNS, load balancing, security, autoscaling, high availability, and distributed-system troubleshooting.
Culture & Benefits
- People-first environment focused on collaboration, accountability, empathy, ownership, and continuous improvement.
- Employment is provided exclusively through an employment contract with FTE status.
- 30-day holidays after five years of tenure, sabbatical leave, and four weeks of paternity leave.
- Annual study and volunteering days, an education budget of up to 8,700 PLN, and access to LinkedIn Learning and English/Polish classes.
- Fully paid medical care, group life insurance, pension contributions, employee assistance, and office benefits.
- Participation in on-call rotations is required to support production reliability and availability.
ΠΡΠ΄ΡΡΠ΅ ΠΎΡΡΠΎΡΠΎΠΆΠ½Ρ: Π΅ΡΠ»ΠΈ ΡΠ°Π±ΠΎΡΠΎΠ΄Π°ΡΠ΅Π»Ρ ΠΏΡΠΎΡΠΈΡ Π²ΠΎΠΉΡΠΈ Π² ΠΈΡ ΡΠΈΡΡΠ΅ΠΌΡ, ΠΈΡΠΏΠΎΠ»ΡΠ·ΡΡ iCloud/Google, ΠΏΡΠΈΡΠ»Π°ΡΡ ΠΊΠΎΠ΄/ΠΏΠ°ΡΠΎΠ»Ρ, Π·Π°ΠΏΡΡΡΠΈΡΡ ΠΊΠΎΠ΄/ΠΠ, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡΠ΅ ΡΡΠΎΠ³ΠΎ - ΡΡΠΎ ΠΌΠΎΡΠ΅Π½Π½ΠΈΠΊΠΈ. ΠΠ±ΡΠ·Π°ΡΠ΅Π»ΡΠ½ΠΎ ΠΆΠΌΠΈΡΠ΅ "ΠΠΎΠΆΠ°Π»ΠΎΠ²Π°ΡΡΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡΠΈΡΠ΅ Π² ΠΏΠΎΠ΄Π΄Π΅ΡΠΆΠΊΡ. ΠΠΎΠ΄ΡΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β
ΠΠΎΡ ΠΎΠΆΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
Affirm
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (SRE & Platform Reliability)
308Β 000 - 428Β 000PLN
Arista Networks
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Site Reliability Engineer (DevOps)
2 ΡΠ°ΡΠ° Π½Π°Π·Π°Π΄
Senior Platform SRE (AWS)
19 ΡΠ°ΡΠΎΠ² Π½Π°Π·Π°Π΄
Principal Site Reliability Engineer (AWS)
220Β 000 - 280Β 000PLN
3 Π΄Π½Ρ Π½Π°Π·Π°Π΄
DevOps Engineer (AWS)
20 ΡΠ°ΡΠΎΠ² Π½Π°Π·Π°Π΄