Назад
5 дней назад

Senior Site Reliability Engineer (Compute)

196 750 - 243 290$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (Compute) (Kubernetes/private cloud): Building and operating Roblox’s private cloud and Kubernetes-based compute infrastructure with an accent on fault tolerance, automation, observability, and production readiness. Focus on designing resilient cluster-management systems, creating load-testing and performance-monitoring tooling, and standardizing reliability practices across the Compute team.

Location: San Mateo, California, United States. Office-based roles require onsite presence Tuesday through Thursday, with optional presence on Monday and Friday.

Salary: $196,750–$243,290 USD annual base pay.

Company

Roblox builds tools and infrastructure for 3D immersive digital experiences created by a global community of developers and creators.

What you will do

  • Design and develop fault-tolerant systems and libraries for Roblox’s cell infrastructure and cluster lifecycle management.
  • Build and productionize Kubernetes-based private cloud infrastructure.
  • Establish reliability practices, technical reviews, and operational guidance across the Infrastructure Compute group.
  • Automate and standardize platform tooling and support processes.
  • Create load-testing, release-capacity, performance-monitoring, canarying, and alerting services.
  • Analyze systems and designs for production readiness and identify platform capacity or degradation issues.

Requirements

  • Bachelor’s degree in Computer Science or a related engineering field, or equivalent professional experience.
  • At least 6 years of experience as an SRE or Software Engineer.
  • At least 3 years of professional programming experience with Go, Java, or C#.
  • Experience with Kubernetes or similar orchestration systems.
  • Experience building reliable software and tools and driving their adoption across an organization.
  • Ability to work from the United States and attend the San Mateo office onsite Tuesday through Thursday.

Nice to have

  • Experience with Nomad, Vault, and Consul.
  • Experience with large project lifecycles, sprint planning, and breaking complex work into measurable deliverables.

Culture & Benefits

  • Collaborative work across infrastructure and tooling partners.
  • Full-time employment with eligibility for equity compensation and employee benefits.
  • Work focused on reliability, problem-solving, and safer shared digital experiences.
  • Future H-1B sponsorship may not be supported, and certain U.S. visa categories may not be eligible for employment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →