Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Kubernetes/GCP): Building and maintaining resilient infrastructure for a software creation platform serving millions of developers with an accent on observability, automation, and high availability. Focus on designing self-healing systems, leading incident response, defining SLOs and SLIs, and optimizing performance across global regions.
Location: Remote - US
Base salary: $210,000–$275,000 per year, plus equity.
Company
Replit is an agentic software creation platform that enables people to build applications using natural language.
What you will do
- Design monitoring, alerting, logging, dashboards, and metrics to improve system observability.
- Build infrastructure automation and infrastructure-as-code solutions with Terraform, Ansible, Pulumi, and related tools.
- Design and maintain reliable CI/CD pipelines and self-healing systems.
- Define and track SLOs and SLIs with product and engineering teams.
- Lead incident response, conduct post-mortems, maintain runbooks, and reduce MTTR.
- Identify infrastructure bottlenecks, plan capacity, optimize resource utilization, and reduce latency across global regions.
Requirements
- 4–8 years of experience in Site Reliability Engineering, DevOps, Systems Engineering, or Infrastructure Engineering.
- Strong programming skills in Python, Go, or a similar automation language.
- Deep understanding of distributed systems, Kubernetes, and cloud-native technologies.
- Experience implementing and maintaining monitoring and observability solutions.
- Experience with infrastructure as code, configuration management, and incident response leadership.
- Experience with Google Cloud Platform and observability tools such as Prometheus, Grafana, or Datadog.
Culture & Benefits
- Full-time employee position with salary and equity.
- 401(k) program with a 4% match, US health, dental, vision, and life insurance.
- Paid parental, medical, and caregiver leave, flexible time off, and holidays.
- Autonomous work environment and monthly wellness stipend.
- Quarterly team gatherings and eligible commuter, office setup, and office amenity benefits.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Site Reliability Engineer (AWS)
120 000 - 185 000$
4 дня назад
Sr. Site Reliability Engineer
160 000 - 180 000$
5 дней назад
Site Reliability Engineer (Kubernetes)
180 000 - 220 000$
6 дней назад
Site Reliability Engineer - Vice President (Kubernetes)
130 000 - 160 000$
8 дней назад
Senior Staff Site Reliability Engineer
232 338 - 290 422$
3 дня назад