11 дней назад
Principal Site Reliability Engineer (Paze)
194 000 - 237 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Site Reliability Engineer (Paze) (SRE/Cloud/DevOps): Improving the reliability, resilience, scalability, observability, and operational health of production services with an accent on distributed systems, automation, cloud infrastructure, and incident response. Focus on defining service-health objectives, reducing operational toil, leading critical incident response, and establishing enterprise-wide reliability practices.
Location: Scottsdale, with a hybrid work model; positions may also be located in Phoenix, Chicago, Washington, New York, or San Francisco. Candidates must be eligible to work in the United States.
Salary: $194,000–$237,000 per year in Phoenix, Chicago, and Washington; $232,000–$284,000 per year in New York and San Francisco, plus discretionary incentives and benefits.
Company
provides payment, financial access, and transaction-protection solutions, including Paze.
What you will do
- Improve how production services are built, tested, deployed, observed, operated, and recovered through software engineering, automation, and DevOps practices.
- Define and improve SLIs, SLOs, error budgets, monitoring, alerting, dashboards, tracing, and other service-health measures.
- Drive continuous improvement across CI/CD, Infrastructure as Code, testing, deployment practices, capacity management, resilience, and operational readiness.
- Partner with Software Engineering and technology teams to embed reliability, scalability, performance, recoverability, and observability throughout the development lifecycle.
- Lead or participate in incident response, troubleshooting, service restoration, post-incident learning, and sustainable on-call operations.
- Establish enterprise technical direction, reduce operational toil, and create reusable automation, tooling, patterns, and engineering practices.
Requirements
- Principal-level expertise in software engineering, systems engineering, Site Reliability Engineering, cloud or platform engineering, DevOps, infrastructure, or a comparable discipline.
- Typically 15+ years of relevant professional experience.
- Experience with software development or scripting, distributed systems, production troubleshooting, automation, and observability.
- Experience with public cloud technologies, preferably AWS, plus infrastructure, networking, Linux/Unix, and modern application architectures.
- Strong analytical, problem-solving, communication, collaboration, and technical leadership skills.
- Must independently possess eligibility to work in the United States; visa sponsorship is not available.
Nice to have
- Experience with AWS or another major cloud platform such as Azure, GCP, or OCI.
- Experience operating highly available production software or distributed systems.
- Experience with containers, orchestration, observability, monitoring, alerting, CI/CD, Infrastructure as Code, and software-delivery automation.
- Experience with incident management, performance and capacity analysis, resilience testing, disaster recovery, or operational readiness.
- Bachelor's degree in a relevant technical field or equivalent practical experience.
Culture & Benefits
- Hybrid work environment focused on collaboration.
- Medical, dental, and vision coverage with HSA and FSA options.
- 401(k) plan with a 100% company safe-harbor match on the first 6% deferred.
- Flexible time off or paid time off, 11 company holidays, and a paid volunteer day.
- 12 weeks of paid parental leave and family-planning support.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 дней назад
Senior Site Reliability Engineer (Temporal)
180 000 - 200 000$
12 дней назад
Staff Site Reliability Engineer (AI)
252 000 - 308 000$
13 дней назад
Site Reliability Engineer (AI)
106 029 - 118 503$
11 дней назад
Site Reliability Engineer (AWS)
120 000 - 185 000$
11 дней назад
Senior Site Reliability Engineer (Cybersecurity)
112 700 - 177 100$
11 дней назад
Senior Infrastructure SRE
139 000 - 155 000CAD