12 дней назад
Senior Staff Engineer, Site Reliability Engineering
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Staff Engineer, Site Reliability Engineering (SRE/Cloud Infrastructure): Driving global reliability strategy and operational excellence for Aiven’s multi-cloud platform with an accent on resilient architecture, incident management, and metrics-driven operations. Focus on designing follow-the-sun on-call models, defining SLIs/SLOs and error budgets, and building automation that improves reliability across 24/7/365 production systems.
Location: Hybrid work policy; the posting does not specify a country or city.
Company
provides a cloud data platform built around open-source databases, search, streaming, application infrastructure, and multi-cloud solutions.
What you will do
- Define and drive the global SRE operating strategy with regional SRE leaders across EMEA, AMER, and APAC.
- Provide hands-on technical leadership across multi-regional SRE and engineering teams.
- Set the reliability engineering vision and roadmap, including tools, automation, and process improvements.
- Improve global incident management, on-call coverage, escalation frameworks, and high-availability operations.
- Establish a metrics-driven operating cadence using KPIs, SLIs, SLOs, and error budgets.
Requirements
- Senior-level experience in SRE or infrastructure engineering across multiple regions or time zones.
- Experience defining and executing reliability strategies at scale, including SLIs, SLOs, incident management, and operational excellence programs.
- Experience operating in a 24/7/365 production environment, with knowledge of follow-the-sun models, on-call design, and large-scale incident response.
- Strong technical leadership, mentoring, and cross-functional influencing skills.
- Technical foundation in distributed systems, cloud infrastructure, and automation.
- Experience driving company-wide technical and operational change and improving efficiency and reliability.
Culture & Benefits
- Work with a global team of more than 400 people across a 24/7/365 operating environment.
- Hybrid work policy with equipment choice.
- Participation in an equity plan.
- Professional Development Plan budget for learning opportunities.
- Global Employee Assistance Program and country-specific benefits.
- Global time-off commitment including parental leave, sick leave, and personal time.
Hiring process
- Submit a resume for consideration.
- Interviews are conducted by video call or in person.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
12 дней назад
Manager, Site Reliability Engineering (Cloud Infrastructure)
5 дней назад
Site Reliability Engineering Team Lead (Principal SRE, Automotive AI)
132 000 - 211 400$
13 дней назад
Senior Site Reliability Engineer (AI)
13 дней назад
Associate Site Reliability Engineer
13 дней назад
Senior Site Reliability Engineer
13 дней назад