18 часов назад
Senior Site Reliability Engineer (Compute)
177 688 - 341 734$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Compute) (SRE/Cloud Systems): Building and operating large-scale, massively distributed, and fault-tolerant systems with an accent on automation, scalability, monitoring, and incident response. Focus on designing robust infrastructure, diagnosing production issues, defining SLOs, SLIs, and SLAs, and addressing performance bottlenecks.
Location: Seattle R&D; fully in-person schedule up to 5 days a week
Salary: $177,688–$341,734 annually, with potential additional bonuses, incentives, and restricted stock units.
Company
operates data privacy and cybersecurity safeguards for U.S. user data, applications, algorithms, and the content ecosystem.
What you will do
- Develop and maintain automation procedures that improve system efficiency and reduce manual intervention.
- Design, deploy, and operate robust systems in collaboration with software engineering teams.
- Build for scalability across web traffic, data growth, and large-scale distributed workloads.
- Implement monitoring, metrics, and performance tests to track system health and resolve bottlenecks.
- Participate in on-call rotations, incident management, diagnosis, remediation, and blameless postmortems.
- Collaborate across teams to define SLOs, SLIs, and SLAs and provide sustainable user support.
Requirements
- Bachelor’s degree in Computer Science, Information Technology, or a related field and 3+ years of experience.
- Professional experience as a Site Reliability Engineer, Systems Engineer, or similar software engineering professional.
- Programming experience with Go, Python, or other languages, with a focus on automation and operational excellence.
- Experience with network architecture, database modeling, cloud systems, and large-scale distributed systems.
- Strong knowledge of Linux, open-source technologies, complex debugging, problem-solving, and cross-functional collaboration.
- Ability to work fully in person in Seattle, with on-site presence up to five days per week.
Nice to have
- Knowledge of monitoring tools and methodologies such as Prometheus and Grafana.
- Experience with containers and orchestration platforms such as Docker and Kubernetes.
Culture & Benefits
- Medical, dental, and vision insurance from day one.
- 401(k) savings plan with company match, paid parental leave, disability coverage, and life insurance.
- Wellbeing benefits, 10 paid holidays, 10 paid sick days, and 17 days of paid personal time.
- Collaborative, diverse, intellectually curious, and problem-solving-oriented work environment.
- Reasonable accommodations are available during recruitment for eligible candidates.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
19 часов назад
Site Reliability Engineer, Tech Infra - USDS (Cloud Infrastructure)
136 800 - 259 200$
9 часов назад
Site Reliability Engineer, Platform Responsibility - USDS (AI/LLM)
129 960 - 246 240$
7 часов назад
Site Reliability Engineer, Edge Services - USDS
136 800 - 359 720$
1 день назад
Staff Site Reliability Engineer (GCP/Kubernetes)
112 500 - 187 500$
5 дней назад
Senior Software Engineer — Observability & IRM (Kubernetes)
124 900 - 250 000$
Okta
7 дней назад
Staff Site Reliability Engineer (Networking)
174 000 - 238 000$