21 час назад
Senior Site Reliability Engineer (Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Kubernetes): Building resilient infrastructure and application reliability solutions for Semrush’s platform with an accent on failure recovery, observability, SLOs, and scalable system architecture. Focus on designing full-stack platform solutions, automating operations with Go and Python, leading critical incidents, and improving service reliability under on-call conditions.
Location: Cyprus
Company
is a brand visibility platform delivering AI-driven insights across SEO, search optimization, content marketing, paid media, and social strategy.
What you will do
- Lead changes to common engineering practices across the company.
- Induce application failures, recover systems, and lead critical incident response.
- Debug applications with metrics and implement tracing and additional observability.
- Establish and refine SLOs, cost dashboards, and security hardening initiatives.
- Collaborate with development teams on scalable, reliable, and efficient system architecture.
- Design full-stack platform solutions and build operational tooling in Go and Python.
Requirements
- 3+ years of experience as a Site Reliability Engineer.
- Experience with Kubernetes and cloud providers.
- Engineering experience with Python or Go.
- Strong understanding of application failures, recovery, metrics-based debugging, traces, and observability.
- Willingness to participate in on-call rotation, typically one week every 2–3 weeks, including possible overnight incidents.
- Strong communication and teamwork skills.
Nice to have
- GCP knowledge.
Culture & Benefits
- Unlimited paid time off.
- Hobby and team-building budget.
- Employee Support Program.
- Financial aid following the loss of a family member.
- Employee Resource Groups.
- Career paths toward Staff, Principal, and engineering management roles.
Hiring process
- Applications are reviewed by the Talent Acquisition team, with feedback typically provided within three working days.
- Interviews cover the candidate’s background, working style, and personality in detail.
- Online interviews use a laptop or desktop, camera, quiet environment, and phone backup for technical issues.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Affirm
6 дней назад
Senior Site Reliability Engineer (SRE & Platform Reliability)
308 000 - 428 000PLN
Cloud.ru
3 дня назад
SRE Engineer
Canonical
6 дней назад
Middle/Senior Site Reliability Engineer (Cloud)
2 дня назад
Senior Site Reliability Engineer (remote within EMEA)
1 день назад
Senior Site Reliability Engineer (Kubernetes)
NDA
5 дней назад
SRE Lead (Fintech)
10 000€