Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Compute) (Kubernetes/private cloud): Building and operating Roblox’s private cloud and Kubernetes-based compute infrastructure with an accent on fault tolerance, automation, observability, and production readiness. Focus on designing resilient cluster-management systems, creating load-testing and performance-monitoring tooling, and standardizing reliability practices across the Compute team.
Location: San Mateo, California, United States. Office-based roles require onsite presence Tuesday through Thursday, with optional presence on Monday and Friday.
Salary: $196,750–$243,290 USD annual base pay.
Company
Roblox builds tools and infrastructure for 3D immersive digital experiences created by a global community of developers and creators.
What you will do
- Design and develop fault-tolerant systems and libraries for Roblox’s cell infrastructure and cluster lifecycle management.
- Build and productionize Kubernetes-based private cloud infrastructure.
- Establish reliability practices, technical reviews, and operational guidance across the Infrastructure Compute group.
- Automate and standardize platform tooling and support processes.
- Create load-testing, release-capacity, performance-monitoring, canarying, and alerting services.
- Analyze systems and designs for production readiness and identify platform capacity or degradation issues.
Requirements
- Bachelor’s degree in Computer Science or a related engineering field, or equivalent professional experience.
- At least 6 years of experience as an SRE or Software Engineer.
- At least 3 years of professional programming experience with Go, Java, or C#.
- Experience with Kubernetes or similar orchestration systems.
- Experience building reliable software and tools and driving their adoption across an organization.
- Ability to work from the United States and attend the San Mateo office onsite Tuesday through Thursday.
Nice to have
- Experience with Nomad, Vault, and Consul.
- Experience with large project lifecycles, sprint planning, and breaking complex work into measurable deliverables.
Culture & Benefits
- Collaborative work across infrastructure and tooling partners.
- Full-time employment with eligibility for equity compensation and employee benefits.
- Work focused on reliability, problem-solving, and safer shared digital experiences.
- Future H-1B sponsorship may not be supported, and certain U.S. visa categories may not be eligible for employment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Software Engineer, Site Reliability
180 000 - 225 000$
Camunda
9 дней назад
Senior Site Reliability Engineer (Kubernetes)
149 800 - 241 500$
8 дней назад
Senior Site Reliability Engineer (Healthcare)
200 000 - 240 000$
11 дней назад
Site Reliability Engineer, Observability
160 000 - 200 000$
6 дней назад
Senior SRE (Kubernetes)
150 000 - 170 000$
9 дней назад