11 дней назад
Senior Site Reliability Engineer - Platform Reliability (Resilience)
71 200 - 92 400€
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer - Platform Reliability (Resilience) (Kubernetes, Terraform, Golang): Designing, building, and scaling Elastic’s multi-cloud platform for hosted and serverless services with an accent on infrastructure automation, Kubernetes operations, and platform resilience. Focus on preventing customer impact during major incidents, improving observability and alerting, and developing reliable systems across distributed cloud environments.
Location: Greece
Salary: €71,200–€92,400 EUR typical starting annual base salary range; no variable compensation.
Company
develops a cloud-based Search AI Platform for search, security, and observability, used by organizations including more than 50% of the Fortune 500.
What you will do
- Design, build, scale, and mature the multi-cloud platform hosting Cloud Hosted and Serverless services.
- Develop software, infrastructure tooling, and automation that support platform engineering and rapid product delivery.
- Lead technical initiatives that automate system engineering and improve the reliability of global infrastructure.
- Operate and improve Kubernetes-at-scale infrastructure across public cloud providers.
- Respond to major incidents, prevent recurring customer impact, and improve problem-management processes.
- Participate in a follow-the-sun on-call rotation, primarily during working hours.
Requirements
- Professional experience in site reliability engineering, software engineering, and Linux system administration on distributed systems at scale.
- Experience operating SaaS products in public cloud environments and using infrastructure-as-code tools such as Terraform or Crossplane.
- Experience building or operating Kubernetes infrastructure at scale, including automation for multi-cloud environments.
- Ability to write non-trivial programs in Golang or another programming language and work with containerized services such as Docker.
- Experience improving alerting, metrics, and major incident management using tools such as Stack, Graphite, Prometheus, or Influx.
- Ability to work from Greece.
Nice to have
- Experience with managed Kubernetes services.
- Experience working in distributed or remote teams.
- Experience designing and implementing solutions with the Stack.
- Experience coaching and mentoring team members.
Culture & Benefits
- Inclusive, collaborative culture focused on operational excellence, continuous improvement, and knowledge sharing.
- Flexible locations and schedules for many roles, with benefits adapted across regions.
- Health coverage for employees and families in many locations.
- Generous vacation allowance and at least 16 weeks of parental leave.
- Up to $2,000 or local-currency equivalent matched for financial donations, plus up to 40 volunteer hours annually.
Hiring process
- Applicants may need to complete role-based security and privacy training.
- Employment may require an export license for individuals located in or nationals of certain sanctioned countries or regions.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →