Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Software Engineer (Observability) (Prometheus/Thanos/Kubernetes): Building and scaling monitoring, logging, and distributed tracing platforms for Reddit with an accent on performance, availability, and event-driven infrastructure. Focus on distributed query systems, large-scale observability data processing, Kubernetes-based systems, and contributing upstream changes to open-source tools.
Location: Remote - United States
Salary: $217,000–$303,900 USD per year, plus equity and applicable commission.
Company
Reddit operates a large-scale community platform with more than 100,000 active communities and approximately 130 million daily active unique visitors.
What you will do
- Build and maintain the foundational observability platform for Reddit's infrastructure.
- Improve the availability, scalability, latency, and efficiency of monitoring, logging, and tracing components.
- Develop and operate large-scale systems based on Prometheus, Thanos, Vector, ClickHouse, OpenTelemetry, and Grafana.
- Automate critical parts of event-driven development and contribute to the technical direction of eventing.
- Participate in on-call rotations and troubleshoot complex systems and software issues.
- Contribute changes upstream to the open-source projects used by the platform.
Requirements
- 7+ years of experience developing internet-scale software, preferably infrastructure.
- Experience with distributed systems development and large-scale system engineering.
- Experience developing on Kubernetes or similar distributed systems.
- Strong troubleshooting skills across systems and software.
- Ability to track work, independently drive projects, and collaborate effectively with a service-oriented team.
- Excellent communication skills.
Nice to have
- Experience with Prometheus, Thanos, Grafana, Vector, ClickHouse, OpenTelemetry, or Loki.
- Kubernetes controller or operator development experience.
Culture & Benefits
- Fully remote work with a flexible-first workforce.
- Comprehensive healthcare and income replacement programs.
- 401(k) with employer match.
- Flexible vacation, paid volunteer time off, and generous paid parental leave.
- Family planning, gender-affirming care, mental health, coaching, professional development, caregiving, and workspace benefits.
Hiring process
- Interviews may be recorded, transcribed, and summarized by AI for select roles; candidates may opt out before scheduled interviews.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Senior DevOps Engineer (AI)
170 000 - 185 000$
Anthropic
3 дня назад
Software Engineer, Infrastructure, Interpretability (AI)
320 000 - 485 000$
5 дней назад
Senior Site Reliability Engineer
CrowdStrike
2 дня назад
Sr. Infrastructure Engineer (Kubernetes)
140 000 - 215 000$
6 дней назад
Sr. Site Reliability Engineer (Kubernetes)
120 000 - 180 000$
4 дня назад
Staff Software Engineer (Platform/DevOps)
160 000 - 208 000$