6 дней назад
Site Reliability Engineer III (DBA)
125 000 - 150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer III (DBA) (Vitess/MySQL and Cassandra): Designing and operating highly available production database systems with an accent on database architecture, disaster recovery, automation, and service reliability. Focus on optimizing distributed databases, leading incident response, building operational runbooks, and improving Kubernetes-based infrastructure.
Location: Remote - US
Salary: $125,000–$150,000 per year
Company
provides open cloud object storage for more than 500,000 customers across 175+ countries and manages over three billion gigabytes of data.
What you will do
- Design, deploy, and own highly available Vitess distributed MySQL and Cassandra database architectures.
- Optimize database performance through query tuning, indexing, schema design, capacity planning, replication, and resharding.
- Own backup, recovery, replication, disaster recovery, security, access control, patching, and hardening practices.
- Monitor production services using SLIs, SLOs, error budgets, logging, alerting, Prometheus, Grafana, Catchpoint, and ELK.
- Develop automation and infrastructure tooling with Python, Bash, Go, Terraform, Ansible, Jenkins, Kubernetes, and Docker.
- Lead incident response, production readiness reviews, root cause analysis, training, documentation, and reliability projects.
Requirements
- 6–8 years of experience in site reliability engineering, systems engineering, infrastructure operations, database engineering, or similar roles.
- Deep hands-on experience with MySQL, distributed or sharded databases, SQL performance analysis, and high-availability architecture.
- Experience administering NoSQL databases such as Cassandra; production experience with Vitess is strongly preferred.
- Strong Linux administration, security operations, incident response, monitoring, alerting, SLI/SLO, and error-budget experience.
- Experience with Kubernetes, Docker, Terraform, Ansible, Jenkins, Vault, Nomad, and at least one of Python, Bash, or Go.
- Experience creating runbooks and operational procedures and mentoring or onboarding engineers; a bachelor’s degree or equivalent professional experience is required.
Nice to have
- Experience in SaaS, cloud services, service provider, or large-scale distributed systems environments.
- Experience with AWS, GCP, Azure, ITIL/OSS practices, and SLA/SLO management.
Culture & Benefits
- Healthcare coverage for families, including dental and vision.
- 401(k), RSU grants for full-time employees, and an ESPP program.
- Flexible vacation policy, maternity and paternity leave, fertility support, and childcare benefits.
- MacBook Pro, workstation stipend, commuter benefits, and learning and development support.
- Work-life balance-focused culture and equal-opportunity employment practices.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Site Reliability Engineer (AI Infrastructure)
175 000 - 265 000$
9 дней назад
Head of Site Reliability Engineering (AI)
195 000 - 285 000$
Baseten
8 дней назад
Site Reliability Engineer (AI)
165 000 - 330 000$
12 дней назад
Senior Cloud Infrastructure and Networking
125 000 - 135 000$
11 дней назад
Principal Site Reliability Engineer (Kubernetes)
190 000 - 220 000$
10 дней назад
DevOps & SRE Engineer (Kubernetes)
100 000 - 150 000$