Назад
4 дня назад

Staff Site Reliability Engineer, Ads

217 000 - 303 900$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff Site Reliability Engineer, Ads (Distributed Systems/AdTech): Leading reliability initiatives and building platforms, tooling, and automation for Reddit's advertising ecosystem with an accent on scalability, operational excellence, and developer productivity. Focus on designing highly available systems, establishing reliability metrics and SLOs, and leading complex incident response across revenue-critical ad-serving, auction, reporting, measurement, and billing systems.

Location: San Francisco, CA; remote-friendly workforce

Salary: $217,000–$303,900 USD per year, plus potential equity and commission depending on the position offered.

Company

Reddit operates a large community platform and an advertising organization serving highly engaged online communities.

What you will do

  • Lead reliability initiatives across ad serving, auctions, targeting, reporting, measurement, and billing.
  • Define a reliability roadmap with engineering leadership covering scalability, operational excellence, and developer productivity.
  • Design and build platforms, tooling, and automation for reliable operation at scale.
  • Lead architecture reviews and influence technical decisions for critical revenue-generating systems.
  • Participate in on-call rotations, investigate complex incidents, and coordinate cross-functional responses to major production events.
  • Establish reliability metrics and mentor engineers across multiple teams.

Requirements

  • 8+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or related roles operating large-scale distributed systems.
  • Experience evolving high-traffic, user-facing production environments and designing highly available systems.
  • Deep expertise in distributed systems, scale engineering, and cloud-native architectures.
  • Strong software engineering skills in general-purpose backend languages such as Go.
  • Understanding of observability systems, including metrics, logging, tracing, and alerting.
  • Experience with SLOs, automation, incident management, performance optimization, and troubleshooting across modern distributed system stacks.

Nice to have

  • Experience supporting advertising technology or other large-scale revenue-critical systems.
  • Experience with ad serving, real-time auctions, budget pacing, campaign delivery, measurement, attribution, or billing systems.
  • Experience operating high-QPS, low-latency services and reliability programs with measurable business outcomes.
  • Experience with Kubernetes, cloud infrastructure, Kafka, ClickHouse, Spark, Flink, BigQuery, or similar platforms.
  • Experience supporting machine learning inference or recommendation systems at scale.

Culture & Benefits

  • Remote-friendly and flexible-first workforce.
  • Health benefits and 401(k) matching for U.S.-based employees.
  • Home-office workspace benefits and personal and professional development funds.
  • Flexible vacation, Reddit Global Days Off, and paid volunteer time off.
  • Family planning support and 4+ months of paid parental leave.

Hiring process

  • Interviews may be recorded, transcribed, and summarized by AI for select roles; opting out is available before scheduled interviews.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →