Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AI): Building reliability tooling and automation for Vapi’s latency-sensitive real-time voice AI infrastructure with an accent on observability, incident response, capacity, and production performance. Focus on debugging across application and infrastructure boundaries, preventing failures in distributed systems, and improving recovery through durable engineering solutions.
Location: San Francisco, hybrid
Salary: $280,000–$314,000 base salary plus equity
Company
Vapi provides a real-time voice AI platform for developers and businesses, powering voice agents and large-scale customer calls.
What you will do
- Build software tooling and automation for core infrastructure spanning compute, storage, networking, and telephony.
- Improve observability, incident response, capacity planning, and production performance.
- Analyze failure modes and deliver durable improvements to failure prevention and recovery.
- Debug across application and infrastructure boundaries in latency-sensitive distributed systems.
- Own a reliability workstream and define the roadmap for future reliability investments.
Requirements
- Senior-level software engineering experience with meaningful SRE, production engineering, or infrastructure experience in distributed systems.
- Experience writing production-quality software and building reliability tooling or automation.
- Deep knowledge of observability, incident response, failure analysis, capacity, and production reliability practices.
- Comfort with Kubernetes, networking, and cloud infrastructure.
- Ability to balance reliability investments with product and engineering velocity.
Nice to have
- Experience with real-time networking or telephony, Envoy, Postgres, Redis, Kafka, Aurora, ClickHouse, or a Google-style SRE environment.
Culture & Benefits
- Ownership-oriented culture with an emphasis on craftsmanship, urgency, direct feedback, and shared accountability.
- Medical, dental, and vision coverage.
- Flexible time off and quarterly team off-sites.
- Catered meals, transportation, gym access, and a $10,000 annual learning and development budget.
- Base salary plus equity ownership.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Staff Site Reliability Engineer (AI)
252 000 - 308 000$
Nscale
8 дней назад
Senior Site Reliability Engineer (AI Infrastructure Operations)
170 000 - 265 000$
Nscale
8 дней назад
Site Reliability Engineer (AI/GPU)
130 000 - 200 000$
9 дней назад
Senior Site Reliability Engineer (AI)
191 000 - 226 000$
7 дней назад
Senior Site Reliability Engineer (Temporal)
180 000 - 200 000$
12 дней назад
Senior Manager, Site Reliability Engineering (AI Ops)
222 000 - 300 500$