4 дня назад
Senior Site Reliability Engineer
140 000 - 180 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (SRE) (Cloud Infrastructure): Building and scaling reliability practices, observability systems, and automation for critical distributed services with an accent on SLOs, incident management, and production-scale cloud infrastructure. Focus on designing actionable alerting, reducing operational toil, leading complex incident response, and improving system resilience through reusable tooling.
Location: Remote, US. Applicants must be legally authorized to work in the United States without current or future employer-sponsored visa sponsorship.
Annual US hiring range: $140,000–$180,000.
Company
provides an AI-powered, cloud-native contact center platform focused on customer experience, automation, multimodality, security, and scalable data insights.
What you will do
- Lead improvements to reliability, scalability, and performance across critical distributed services.
- Define and implement SLIs, SLOs, and error budgets to guide engineering priorities.
- Design observability systems covering metrics, logging, tracing, and actionable alerting.
- Lead complex incident response, act as incident commander when needed, and drive systemic postmortem actions.
- Eliminate operational toil through automation, reusable tooling, improved workflows, and paved roads.
- Partner with product and platform teams on architecture, production readiness, failure recovery, and operational maturity while mentoring engineers.
Requirements
- 6–10+ years of experience in SRE, infrastructure, or backend systems engineering.
- Experience owning reliability outcomes for complex distributed systems and production-scale infrastructure.
- Strong experience with AWS, GCP, or Azure cloud infrastructure.
- Deep understanding of observability, incident management, and system performance.
- Proficiency in at least one programming language such as Go, Python, or Java, with a focus on automation and tooling.
- Ability to influence engineering teams without formal managerial authority and make clear decisions during incidents.
Nice to have
- Experience building or scaling SRE practices, including SLOs, incident frameworks, and on-call models.
- Kubernetes or container orchestration experience.
- Infrastructure as Code experience with Terraform or similar tools.
- Experience with high-growth systems, performance engineering, or capacity planning.
Culture & Benefits
- Work on AI-powered customer experience and contact center technology.
- Collaborative and inclusive environment focused on creative solutions and strong working relationships.
- Medical, dental, and vision coverage.
- 401(k) plan, wellness benefits, and additional benefits.
- Equal employment opportunity and reasonable accommodation support, including an alternative to AI-assisted screening.
Hiring process
- uses AI-assisted technology for initial application review and filtering against the listed qualifications.
- Human recruiting and hiring teams conduct application review decisions, screening, and interviews.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Site Reliability Engineer (AI Agents & Automation)
147 600 - 221 400$
10 дней назад
Principal Site Reliability Engineer (AI)
165 000 - 185 000$
Replit
10 дней назад
Site Reliability Engineer
210 000 - 275 000$
5 дней назад
Senior Site Reliability Developer (SaaS)
123 250 - 166 750CAD
6 дней назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
9 дней назад
Site Reliability Engineering Manager (AWS/Kubernetes)
205 000 - 255 000$