Staff Software Engineer - SRE
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Staff Software Engineer - SRE (SRE/Observability): Own reliability for core production systems, drive capacity planning, and improve observability to reduce incident response time with an accent on distributed systems reliability at scale. Focus on building automation and operational patterns across deployments, on-call, incident response, and database/data-processing safety.
Location: Kraków, Lesser Poland Voivodeship, Poland
Company
builds a cloud-based platform for smart security and IoT.
What you will do
- Own reliability for core production systems, including deployments, on-call operations, and incident response.
- Drive capacity planning across cloud and physical infrastructure.
- Define and improve reliability goals and operational outcomes (e.g., availability and MTTR).
- Lead observability improvements across metrics, logs, traces, alert quality, and runbook maturity.
- Build automation and tooling to improve reliability and engineering velocity.
- Partner with engineering and product teams to align reliability priorities with roadmap goals, including database and data-processing reliability.
Requirements
- 10+ years of professional software engineering experience, including production operations and on-call ownership.
- Proven track record of improving measurable reliability outcomes in distributed systems.
- Deep expertise in observability and production support practices.
- Strong networking fundamentals for distributed systems (TCP/IP, DNS, TLS, HTTP, and L4/L7 behavior).
- Strong experience with distributed systems technologies such as Kubernetes, Kafka, and Redis.
- Strong experience with cloud infrastructure and operations at scale.
Culture & Benefits
- Work fully from the office in Kraków, with the possibility to work remotely occasionally.
- Permanent employment contract (CoE) and long-term employment.
- Medical care, life insurance, sports package, and an annual professional development budget ($2,000).
- Hands-on ownership with a focus on reliability, customer experience, and engineering quality.
Hiring process
- Interviews focused on reliability/production operations, observability, and distributed systems problem-solving.
- Discussion of how reliability goals translate into practical operational outcomes.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →