1 день назад
Senior Site Reliability Engineer (Kubernetes)
93 700 - 138 700$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Kubernetes): Designing and operating scalable multi-cloud and hybrid infrastructure for game services, account platforms, CI/CD pipelines, and developer tooling with an accent on Kubernetes platforms, infrastructure as code, observability, and reliability engineering. Focus on building progressive delivery systems, defining SLI/SLO and error-budget policies, leading chaos engineering and incident response, and automating secure platform operations across AWS, GCP, and on-premises environments.
Location: Burnaby, British Columbia, Canada; hybrid
Salary: $93,700–$138,700 per year at the start of employment, with potential additional bonus and/or equity.
Company
develops and publishes major video game franchises and supports a portfolio of global game development studios.
What you will do
- Design, build, and operate scalable multi-cloud and hybrid infrastructure across AWS, GCP, and on-premises data centers.
- Own Kubernetes platforms on EKS and GKE, including cluster lifecycle, multi-tenancy, networking, storage, and autoscaling.
- Build observability with Prometheus, Grafana, Datadog, and OpenTelemetry; define SLI/SLO and error-budget policies.
- Implement progressive delivery, self-service provisioning, automated remediation, intelligent scaling, and secure CI/CD pipelines.
- Lead chaos engineering, incident response, and post-mortems focused on systemic fixes.
- Promote SRE practices across studios, contribute to architecture decisions, and author engineering RFCs.
Requirements
- 5+ years of experience in SRE, platform engineering, or equivalent production-scale infrastructure work.
- Deep Kubernetes experience in cloud environments, preferably EKS or GKE, including networking, storage, and multi-cluster patterns.
- Strong Terraform and/or Pulumi skills, with hands-on experience in Helm, Terragrunt, and GitOps tooling.
- Experience with AWS, GCP, VMware, bare-metal servers, Ansible, Puppet, and AWS Systems Manager.
- Production-quality coding in Go, Python, or TypeScript, plus strong Linux, TCP/IP, DNS, and TLS debugging skills.
- Must be legally authorized to work in Canada without current or future employer sponsorship.
Nice to have
- Live-service game or large-scale consumer internet experience serving millions of concurrent users.
- Advanced service mesh and Kubernetes networking experience with Istio or Cilium.
- FinOps, cloud certifications, AI and agentic development experience, or experience mentoring SREs.
Culture & Benefits
- Systems-focused post-mortems and an automation-first approach to repetitive work.
- Opportunity to support infrastructure connecting millions of players across global regions.
- Regular employees may receive bonuses, equity awards, and medical, financial, and other benefits.
- is committed to an inclusive workplace and provides reasonable accommodations during hiring and employment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Senior Site Reliability Developer (AWS)
107 000 - 157 300CAD
3 дня назад
Associate Infrastructure Engineer (Kubernetes)
140 000 - 190 000$
4 дня назад
Senior DevOps Developer (AWS)
107 000 - 157 300CAD
2 дня назад
Principal Platform Developer (AWS)
131 000 - 191 400CAD
Baseten
4 дня назад
Software Engineer (Platform)
165 000 - 330 000$
5 дней назад
Senior Machine Learning Operations Developer: AI/ML Platform
153 000 - 224 400CAD