2 дня назад
Site Reliability Engineer III (AWS/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer III (AWS/Kubernetes): Building and operating reliable multi-tenant SaaS infrastructure and automation for a global cloud platform with an accent on availability, performance, observability, and maintainability. Focus on automating AWS operations, engineering SSO platforms, improving incident management, and developing self-healing systems at scale.
Location: Ireland — Dublin
Company
provides cloud software and analytics for property and casualty insurance companies, including core applications for policy, claims, billing, data management, digital portals, and predictive analytics.
What you will do
- Manage shared multi-tenant infrastructure for resilient SaaS microservices and customer-facing applications.
- Automate AWS deployment and operational processes and develop core infrastructure reliability features.
- Engineer and maintain SAML/OAuth-based single sign-on infrastructure.
- Build observability tooling, metrics, and dashboards for global platform infrastructure.
- Improve incident management through risk mitigation, blameless postmortems, SLO tracking, and self-healing capabilities.
- Document systems, train teammates, and collaborate with engineering teams on reliability and product improvements.
Requirements
- Bachelor’s degree in Computer Science or a related field.
- Software engineering and automation experience with Bash, Python, and/or Go, plus strong Linux expertise.
- Significant AWS experience supporting live production environments, including Java, Apache, and Tomcat systems.
- Experience with Infrastructure as Code and DevOps/GitOps tools such as Terraform, Terragrunt, Git, Flux CD, and TeamCity.
- Hands-on experience with Docker, Helm, Kubernetes/EKS, CNI, Ingress networking, SSO, SAML, and OAuth.
- Experience with observability tools, event streaming or queues, relational databases, troubleshooting, mentoring, and technical communication.
Nice to have
- Production support experience for SaaS platforms.
- Experience with Okta, KubeVela, or Crossplane.
- Experience with Aurora Postgres, Oracle RDS, Kafka, or AWS SQS.
Culture & Benefits
- Collaborative culture built around integrity, rationality, collegiality, curiosity, and innovation.
- Focus on responsible use of AI, data-driven insights, and continuous improvement.
- Opportunity to work on a cloud platform serving hundreds of insurance providers and millions of daily transactions.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
Site Reliability Engineer (AWS)
2 дня назад
Senior Site Reliability Engineer (Kubernetes)
Wheely
1 день назад
Site Reliability Engineer (AWS/Kubernetes)
5 000€
5 часов назад
Senior Site Reliability Engineer (Kubernetes)
1 день назад
Site Reliability Engineer/L3 Support (AWS/Kubernetes)
110 000 - 130 000$
2 дня назад
Site Reliability Engineer (AWS/Kubernetes)
90 000 - 120 000GBP