5 дней назад
Site Reliability Engineer ll (DBA)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer ll (DBA) (Vitess/MySQL and Cassandra): Operating and automating production database systems and critical services with an accent on database performance, observability, incident response, and data durability. Focus on query tuning, replication and recovery procedures, reliability automation, and monitoring in Kubernetes-based environments.
Location: Remote in Argentina, Colombia, Costa Rica, or Mexico
Company
is an open cloud object storage provider managing large-scale data storage for more than 500,000 customers in over 175 countries.
What you will do
- Operate and maintain highly available Vitess distributed MySQL and Cassandra databases using established architectures and runbooks.
- Improve database performance through query tuning, indexing, and schema design.
- Execute backup, recovery, replication, security, and access-control procedures to protect data durability.
- Monitor production services using SLIs, SLOs, error budgets, and observability tools, and participate in on-call rotations and incident response.
- Develop operational automation and contribute to monitoring, logging, alerting, CI/CD, configuration management, and infrastructure-as-code tooling.
- Support capacity planning, disaster recovery exercises, vendor troubleshooting, documentation, and post-incident improvements.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience.
- 2–4 years of experience in site reliability, systems engineering, or database-focused operations.
- Solid Linux administration and troubleshooting skills, with familiarity with monitoring, alerting, incident response, and root cause analysis.
- Proficiency in Python, Bash, or Go, plus SQL and NoSQL database management experience.
- Hands-on MySQL performance tuning, replication, and disaster recovery experience.
- Understanding of containers, Kubernetes, Docker, microservices, and operational best practices.
Nice to have
- Experience with Vitess or other distributed or sharded MySQL systems.
- Experience in SaaS, service provider, or distributed systems environments.
- Familiarity with ITIL/OSS practices, SLOs, SLAs, and cloud platforms such as AWS, GCP, or Azure.
Culture & Benefits
- Work within established procedures and runbooks while contributing to continuous reliability improvements.
- Collaborate with engineering, product, operations, vendors, and service providers.
- Support a reliability-minded engineering culture through documentation and knowledge sharing.
- Inclusive workplace focused on diversity, equity, and belonging.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →