5 дней назад
Principal Site Reliability Engineer (GCP)
151 000 - 244 200$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Site Reliability Engineer (GCP/Observability): Operating and enhancing large-scale GCP infrastructure and observability systems for a cybersecurity SecOps platform with an accent on high-cardinality metrics, tracing, and large-scale logging. Focus on automating monitoring and remediation, managing incidents and alerts, and improving reliability and availability through Kubernetes, Docker, Python, and infrastructure as code.
Location: Santa Clara, California, United States; most teams work from the office full time.
Annual base salary: $151,000–$244,200.
Company
Cybersecurity organization building a SecOps platform with XDR, XSIAM, XSOAR, and XPANSE products.
What you will do
- Operate and maintain large-scale GCP infrastructure for the Cortex Observability team.
- Design, implement, and continuously improve observability, monitoring, tracing, metrics, and logging systems.
- Develop monitoring, alerting, automated remediation, and auto-scaling tools.
- Manage incident response and alerts to minimize service impact.
- Provide follow-the-sun operational coverage and troubleshoot complex production issues.
- Collaborate with engineering teams to improve product operability, reliability, and availability.
Requirements
- 5+ years of DevOps or SRE experience with a focus on service reliability.
- High proficiency with Prometheus, Grafana, OpenTelemetry, PagerDuty, and Prometheus Alertmanager.
- Strong experience with GCP or AWS, Kubernetes, and Docker.
- Strong Python and Linux Shell skills, plus experience with Ansible and Terraform.
- Ability to coordinate across multiple teams and time zones, troubleshoot complex problems, and work independently.
Culture & Benefits
- In-person collaboration is emphasized, with flexibility when needed.
- Employee benefits are available as part of the compensation package.
- Restricted stock units and a bonus may be included in the offer.
- Immigration sponsorship is available.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
Principal Site Reliability Engineer (Kubernetes)
190 000 - 220 000$
9 дней назад
Head of Site Reliability Engineering (AI)
195 000 - 285 000$
10 дней назад
Site Reliability Engineer Staff (Cloud Infrastructure)
5 дней назад
Site Reliability Engineer (Identity & Access Management)
85 000 - 122 200$
6 дней назад
Site Reliability Engineering Team Lead (Principal SRE, Automotive AI)
132 000 - 211 400$
10 дней назад
Staff Site Reliability Engineer (Cybersecurity)
199 750 - 270 000$