14 часов назад
Senior Manager, Site Reliability Engineering (Application) EMEA (AWS/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Manager, Site Reliability Engineering (Application) EMEA (AWS/Kubernetes): Leading an EMEA SRE team and managing the reliability, performance, and resilience of Guidewire Cloud Platform applications and services with an accent on observability, incident management, and scalable infrastructure. Focus on building reliable cloud systems, solving complex distributed-systems problems, and driving automation across AWS, Kubernetes, Terraform, and GitOps environments.
Location: Krakow, Poland (in office)
Company
provides a cloud platform and software products for property and casualty insurance companies, including core insurance applications, digital portals, data management, predictive analytics, and AI capabilities.
What you will do
- Lead and mentor an EMEA SRE team of 10–17 engineers, including senior individual contributors and technical leads.
- Set operational goals and guide adoption of SRE practices such as observability, reliability engineering, service-level objectives, error budgets, and Production Readiness Reviews.
- Lead strategy for large cross-functional projects and solve complex, loosely defined reliability problems across distributed systems.
- Oversee regional incident management, serve as an escalation point, and ensure blameless postmortems produce preventative actions.
- Manage resources, schedules, stakeholder alignment, and department workstreams to meet business goals.
- Promote automation and improve the reliability and efficiency of Cloud Platform operations.
Requirements
- Leadership experience in Site Reliability Engineering, Software Engineering, or a related field, including management of teams of 10 or more engineers and team leads.
- Typically 10–12 years of related experience with a Bachelor's degree, 8 years with a Master's degree, 5 years with a PhD, or equivalent experience.
- Hands-on software engineering and architectural experience with Go, Python, or Java, plus experience building and supporting microservices with Spring Framework and React.
- Experience implementing SRE principles, managing major incidents, and serving as or working closely with an Incident Commander.
- Deep expertise in AWS, Kubernetes, cloud networking, security, RBAC, vulnerability management, Terraform, GitOps, and CI/CD pipelines.
- Experience with observability and APM tools, people management in Workday, and workstream management using tools such as JIRA Service Management and AHA.
Nice to have
- AWS or Kubernetes certifications such as CKA or CKAD.
- Familiarity with InsuranceSuite, Cloud Platform, or Gosu.
- Experience with Kubevela or applying agentic AI and LLMs to SRE practices.
- Experience with Coupa, Concur, or Salesforce.
Culture & Benefits
- Work with cloud technology supporting more than 540 insurers in 40 countries.
- Culture centered on integrity, rationality, collegiality, innovation, teamwork, and work-life balance.
- Competitive compensation, comprehensive benefits, and career development opportunities.
- No formal on-call rotation for this senior leadership role, but availability is required outside standard business hours, including evenings, weekends, and holidays, for major escalations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
B2B Systems Site Reliability Engineer (AI)
4 дня назад
Senior Engineering Manager (Identity Governance)
7 дней назад
Senior Site Reliability Engineer (Kubernetes, B2B)
2 дня назад
Head of Data Practice
8 600 - 9 900€
GlobalLogic
3 дня назад
Site Reliability Engineer (Senior SRE / Systems DevOps)
6 дней назад