6 часов назад
Staff Site Reliability Engineer (AWS/Kubernetes)
140 000 - 155 000CAD
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Site Reliability Engineer (AWS/Kubernetes): Building and evolving foundational platform systems, secure CI/CD pipelines, and operational tooling for reliable distributed services with an accent on AWS, Kubernetes, observability, and resilience engineering. Focus on defining SLOs, improving incident response, optimizing cloud-native infrastructure, and reducing operational toil through automation.
Location: Remote; candidates must be legally authorized to work in Canada.
Salary: CAD 140,000–155,000 annual base salary, plus discretionary bonus and/or commission.
Company
develops accounting and financial reporting software solutions.
What you will do
- Drive reliability engineering and operational excellence for mission-critical services running on AWS and Kubernetes.
- Design and improve deployment, release, rollback, and secure-by-default CI/CD strategies for distributed systems.
- Build observability capabilities across metrics, logs, tracing, alerting, SLIs, SLOs, and reliability standards.
- Lead high-severity incident response, post-incident reviews, and continuous reliability improvements.
- Partner with Engineering, Security, Platform, and Product teams to improve resilience, performance, and platform standards.
- Mentor engineers in cloud-native technologies, SRE principles, and operational excellence.
Requirements
- 8+ years of experience in SRE, Platform Engineering, DevOps, or related cloud-native engineering roles.
- Deep AWS expertise, including EKS, IAM, VPC, Lambda, CloudFront, S3, and cloud networking and security.
- Advanced experience operating and scaling production Kubernetes environments, including Kubernetes architecture, controllers, Gateway API, ingress, and service networking.
- Strong experience with Istio, Infrastructure as Code using AWS CDK, and CI/CD platforms such as GitHub Actions.
- Strong proficiency in TypeScript and Node.js for platform engineering, automation, and operational tooling.
- Must be legally authorized to work in Canada.
Nice to have
- Experience with canary, blue/green, and feature-flag-based progressive delivery.
- Experience in regulated, compliance-driven, or security-sensitive SaaS environments.
- FinOps and cloud cost optimization experience.
- Experience building internal developer platforms and self-service tooling.
- Cloud-native certifications such as CKA, CKAD, CKS, KCSA, KCNA, or Kubestronaut certification.
Culture & Benefits
- Flexible remote work options and generous time-off policies.
- Health insurance and retirement plans.
- Collaborative, inclusive culture with knowledge sharing.
- Performance bonuses, recognition programs, and career growth opportunities.
- International projects and collaboration with a diverse global team.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Site Reliability Engineer
124 000 - 170 500$
5 дней назад
Staff DevOps Engineer (AWS)
125 000 - 150 000$
6 дней назад
Senior Platform Engineer (AWS/Kubernetes)
82 000 - 115 000€
Latitude
5 дней назад
Senior Site Reliability Engineer (Kubernetes)
6 дней назад
Associate Infrastructure Engineer (Kubernetes)
140 000 - 190 000$
Deimos
17 часов назад