1 день назад
Senior Site Reliability Engineer (SRE)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (SRE) (Node.js/TypeScript): Building reliable and observable systems for a consumer mobile dating product with an accent on critical user journeys, production monitoring, and incident response. Focus on defining SLIs/SLOs, improving metrics, logging, tracing, and alerting, and independently mitigating P0/P1 incidents across AWS-based systems.
Location: Brazil — Remote
Company
connects talent with product careers at high-growth companies and places engineers directly within client product teams through its Extended Team model.
What you will do
- Own observability for critical product and user journeys within an autonomous product squad.
- Define and maintain metrics, dashboards, alerts, SLIs, SLOs, and product-level reliability metrics.
- Improve monitoring, logging, tracing, and alerting across backend and infrastructure systems.
- Act as the first responder for P0/P1 production incidents, including out-of-hours incidents.
- Investigate production signals, identify root causes, mitigate issues, and coordinate escalations.
- Drive reliability improvements through incident triage, postmortems, infrastructure, and tooling enhancements.
Requirements
- Senior hands-on backend or software engineering experience with Node.js and TypeScript.
- Experience in SRE, Production Engineering, or a similar reliability-focused role.
- Strong production experience with AWS, plus experience with Cloudflare and CloudWatch.
- Practical expertise in metrics, logging, tracing, alerting, incident triage, mitigation, and postmortems.
- Ability to understand both application and infrastructure layers and investigate production issues independently.
- Comfort participating in out-of-hours incident response within a distributed engineering team.
Nice to have
- Experience with Sentry and defining SLIs/SLOs for product-level metrics.
- React Native experience or exposure to mobile application environments.
- Experience with consumer mobile products or high-traffic B2C systems.
Culture & Benefits
- Work directly with the client and use the same tools and technologies as the client team.
- Join small, cross-functional product squads distributed across Europe, the US, and LATAM.
- Health insurance, wellbeing budget, sport coverage, and learning budget.
- 18 business days of paid vacation per year, paid sick leave, and paid public holidays.
- Long-term projects, client trips, teambuilding, and after-work activities.
Hiring process
- Recruiter interview with .
- Technical interview.
- Final interview and reference check.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
2 дня назад
Site Reliability Engineer (Node.js/TypeScript)
2 дня назад
Site Reliability Engineer (TypeScript)
180 000 - 220 000$
3 дня назад
Staff Site Reliability Engineer (AWS/Kubernetes)
140 000 - 155 000CAD
Deimos
4 дня назад
Senior Site Reliability Engineer
2 дня назад
Senior Site Reliability Engineer (SRE)
2 дня назад
Senior Site Reliability Engineer (AWS)
140 000 - 180 000$