11 дней назад
Senior Site Reliability Engineer - Platform Reliability (Resilience)
76 000 - 101 800€
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer - Platform Reliability (Resilience) (SaaS/Cloud Infrastructure): Designing, building, scaling, and maturing Elastic’s multi-cloud platform for hosted and serverless services with an accent on reliability automation, Kubernetes infrastructure, and operational excellence. Focus on leading incident and problem management, developing Go-based tooling, and preventing repeated customer impact across globally distributed systems.
Location: Spain
Base salary: €76,000–€101,800 per year; no variable compensation component.
Company
develops the Search AI Platform for search, security, observability, and cloud-based data services used by organizations worldwide.
What you will do
- Design, build, scale, and mature the multi-cloud platform supporting Cloud Hosted and Serverless services.
- Develop and maintain software, infrastructure tooling, and automation for global platform reliability.
- Lead technical initiatives that automate system engineering and improve the reliability of infrastructure.
- Respond to major incidents, improve problem management, and prevent repeated customer impact.
- Participate in a follow-the-sun on-call rotation, primarily during working hours.
- Collaborate with distributed engineering teams while coaching and uplifting colleagues.
Requirements
- Experience operating SaaS products in public cloud environments using Infrastructure-as-Code tools such as Crossplane or Terraform.
- Experience building or operating Kubernetes infrastructure at scale across multiple cloud providers.
- Professional programming experience in Golang or another programming language, plus experience with containerized services such as Docker.
- Experience improving alerting, incident management, metrics, and operational processes using tools such as Stack, Graphite, Prometheus, or Influx.
- Professional Linux system administration experience on distributed systems at scale.
- Experience designing and implementing solutions with the Stack and working effectively in globally distributed teams.
Nice to have
- Experience with public cloud and managed Kubernetes services.
- Experience working remotely or in distributed teams.
Culture & Benefits
- Inclusive, collaborative environment focused on operational excellence, progress, and knowledge sharing.
- Flexible locations and schedules for many roles, with benefits adapted across regions.
- Health coverage for employees and families in many locations.
- Generous vacation allowance and a minimum of 16 weeks of parental leave.
- Up to $2,000 matched annually for financial donations and up to 40 hours for volunteer projects.
- Responsibilities include protecting organizational data and systems and following ’s Secure Software Development Framework.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →