обновлено 3 дня назад
Senior Site Reliability Engineer (MongoDB)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (MongoDB): Operating and improving a large-scale MongoDB platform for real-time customer engagement with an accent on reliability, observability, performance, and automation. Focus on Kubernetes-based cluster lifecycle management, sharding and capacity planning, automated recovery workflows, and systemic incident reduction.
Location: São Paulo, Brazil; hybrid work
Company
provides a customer engagement platform that helps brands deliver personalized experiences through cross-channel messaging, journey orchestration, and AI-powered decisioning.
What you will do
- Design and operate large-scale MongoDB infrastructure with strong availability, durability, SLA, and query-performance requirements.
- Build MongoDB-specific monitoring, alerting, capacity planning, sharding strategies, and observability.
- Partner with product engineering teams on schema design, indexes, aggregation pipelines, connection pools, and read/write standards.
- Manage MongoDB clusters on Kubernetes using the MongoDB Enterprise Kubernetes Operator, Terraform, and Ansible.
- Develop automated backup, restore, point-in-time recovery, and self-service tooling in Python, Go, Ruby, or JavaScript.
- Participate in PagerDuty on-call rotations, lead incident retrospectives, and improve runbooks through systemic fixes and automation.
Requirements
- 5+ years of production experience as a Software Engineer, DevOps Engineer, or Site Reliability Engineer.
- Hands-on MongoDB expertise covering replica sets, sharding, index design, aggregation pipelines, explain plans, and performance tuning under real load.
- Strong Linux fundamentals, including disk I/O, memory, networking, and process management.
- Strong programming skills in Python, Go, Ruby, or JavaScript, with experience writing automation.
- Experience with Terraform or Ansible and with Docker and Kubernetes.
- Systems thinking, attention to failure modes and cascading effects, and a documentation-first approach to asynchronous collaboration.
Nice to have
- Experience operating MongoDB at multi-terabyte scale or in a sharded topology.
- Familiarity with MongoDB Atlas, Ops Manager, or Cloud Manager.
- Experience with Redis, Kafka, or Postgres.
- Prior database platform engineering or database reliability engineering experience.
Culture & Benefits
- Hybrid ways of working with a curated in-office experience.
- Competitive compensation that may include equity, retirement plans, and an employee stock purchase plan.
- Flexible paid time off and comprehensive medical, dental, vision, life, and disability benefits.
- Fertility benefits, equal paid parental leave, professional development support, and a yearly learning stipend.
- Collaborative, transparent culture with employee resource groups, volunteer opportunities, and donation matching.
Hiring process
- AI-assisted tools may be used to analyze application materials and support interview scheduling, recording, and note summarization.
- Recruiting teams remain responsible for hiring decisions; candidates may have location-dependent rights to request information, opt out, or request manual review.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Senior Site Reliability Engineer (AWS/Kubernetes)
5 дней назад
Senior Site Reliability Engineer (Fintech)
4 дня назад
Site Reliability Engineer III, GWCP (SaaS)
4 дня назад
Senior Site Reliability Engineer, GovCloud (AWS)
128 500 - 190 000$
3 дня назад
Site Reliability Engineer (AI)
4 дня назад