9 дней назад
Principal Site Reliability Engineer, Platform Engineering
223 200 - 380 400$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Site Reliability Engineer, Platform Engineering (SaaS/platform engineering): Shaping the architecture and operational model for a fully managed single-tenant SaaS platform with an accent on reliability, security, compliance, tenant orchestration, and scalable platform patterns. Focus on designing resilient cell-based architectures, reducing operational toil through automation, and leading complex technical decisions across teams.
Location: Remote in Canada, the United Kingdom, or the United States
Base salary: $223,200–$380,400 USD per year for US residents
Company
An intelligent orchestration platform for DevSecOps that helps organizations improve developer productivity, operational efficiency, security, compliance, and digital transformation.
What you will do
- Set technical direction for a fully managed, single-tenant SaaS platform and its fleet of isolated customer environments.
- Lead transformations across resilience, failover, tenant orchestration, change management, self-service tooling, and platform integrations.
- Drive modular, cell-based architecture while preserving security, environment isolation, and compliance.
- Strengthen service ownership and operational maturity across engineering teams.
- Identify systemic reliability and scalability risks using production signals, incident patterns, and architectural analysis.
- Establish reusable automation and platform patterns, make cross-team technical decisions, mentor senior engineers, and advance engineering excellence.
Requirements
- Deep expertise in site reliability, platform, infrastructure, or backend engineering with experience operating large-scale production systems.
- Hands-on experience with cloud infrastructure, automation, observability, infrastructure as code, and modern production engineering practices.
- Strong software engineering fundamentals and experience building production systems or infrastructure tooling with Go, Ruby, Python, or similar languages.
- Expertise in distributed systems and systems design, including reliability, failure isolation, scalability, and operational complexity.
- Technical leadership experience across multiple teams, including platform or infrastructure transformations and improvements to production ownership.
- Exceptional technical communication, influence, mentoring, and architectural decision-making skills.
Culture & Benefits
- Fully remote work with country-specific eligibility requirements.
- AI is used as a core productivity multiplier in daily workflows.
- Flexible paid time off and parental leave.
- Health, financial, and well-being benefits.
- Equity compensation, employee stock purchase plan, and a growth and development fund.
- Team member resource groups and a culture focused on inclusion, knowledge exchange, and continuous development.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
14 дней назад
Site Reliability Engineer (AI Platform)
69 768 - 96 900€
12 дней назад
Senior Site Reliability Engineer (Satellite Operations)
142 800 - 178 500$
14 дней назад
Senior Site Reliability Engineer, PaaS (Kubernetes)
69 768 - 96 900€
10 дней назад
Manager, Site Reliability Engineering (Cloud Infrastructure)
11 дней назад
Staff Site Reliability Expert (AWS)
9 дней назад
Site Reliability Engineer (AI)
200 000 - 400 000$