6 дней назад
Site Reliability Engineer III (SRE) - Guidewire Cloud Platform (Application)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer III (SRE) - Guidewire Cloud Platform (Application) (cloud reliability): Ensuring the reliability, performance, and scalability of applications running on the Guidewire Cloud Platform with an accent on automation, observability, and distributed-systems troubleshooting. Focus on designing SLI/SLO/error-budget practices, building automated runbooks and monitoring, and deploying scalable AWS and Kubernetes infrastructure.
Location: Krakow, Poland
Company
provides cloud software and data, digital, analytics, and AI products used by property and casualty insurers worldwide.
What you will do
- Troubleshoot application and distributed-system issues with development teams while minimizing customer impact.
- Develop automated runbooks, monitors, dashboards, and synthetic transactions to reduce manual operational work.
- Improve application reliability, performance, scalability, and service availability on the Cloud Platform.
- Design and maintain SLI, SLO, and Error Budget practices, including application performance monitoring and telemetry.
- Deploy and manage scalable AWS and Kubernetes infrastructure using Terraform and cloud-native approaches.
- Participate in mandatory on-call rotations, including responses outside regular business hours, on weekends, and during holidays.
Requirements
- Experience in SRE or a similar reliability-focused engineering role.
- Software engineering experience with Python, Go, or Java, including clean, testable code and SOLID, DRY, and KISS principles.
- Experience troubleshooting distributed systems on cloud infrastructure and working with CI/CD pipelines in Kubernetes and legacy environments.
- Experience with AWS, Kubernetes, Terraform, configuration management, GitOps, Puppet, or Ansible.
- Knowledge of monitoring, APM, telemetry, cloud networking, security, vulnerability management, and programmatic infrastructure remediation.
- Willingness and ability to participate in mandatory on-call rotations.
Nice to have
- SRE or AWS certification.
- Experience with SQL, database administration, data pipelines, performance tuning, and schema design.
- Familiarity with TeamCity, Bitbucket Pipelines, Jenkins, or GitHub Actions.
- Experience with Hadoop, Apache Spark, AWS Redshift, microservices, or event-driven architectures.
Culture & Benefits
- Mission-driven work supporting insurance customers during crises and other challenging events.
- Culture focused on innovation, teamwork, continuous learning, and work-life balance.
- Competitive compensation and comprehensive benefits.
- Opportunities for career development and collaboration with experienced technical peers.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
12 дней назад
Staff Observability Engineer (AI)
290 000 - 375 000PLN
7 дней назад
Senior Site Reliability Engineer (Kubernetes, B2B)
8 дней назад
Senior SRE or DevOps Engineer (Networking)
10 дней назад
Lead AI SRE and QA Engineer
9 дней назад
IT Operations Engineer
800 - 950PLN
7 дней назад