5 дней назад
Senior Site Reliability Developer (SaaS)
123 250 - 166 750CAD
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Developer (SaaS): Building and operating scalable, resilient, and automated distributed services for a multi-cloud SaaS platform with an accent on AWS, Azure, infrastructure as code, observability, and incident response. Focus on designing fault-tolerant systems, automating remediation, reducing operational toil, and maintaining high availability across production services.
Location: Canada; flexible work options include full-time in the Toronto office, hybrid work, or remote work within Canada. Preference is given to candidates in the GTA who can attend the Toronto office approximately two days per month.
Base salary: CAD 123,250–166,750 per year.
Company
provides a SaaS platform that helps finance teams plan and grow using data, collaboration, and familiar tools.
What you will do
- Build and operate scalable, resilient, fault-tolerant services across AWS and Azure.
- Develop infrastructure, software platforms, frameworks, runbooks, and standard operating procedures.
- Own and lead projects across the SaaS Technology and Operations team.
- Monitor availability, latency, system health, telemetry, logging, and alerting for live services.
- Drive automated incident response, remediation, post-mortems, and proactive outage prevention.
- Mentor SREs and Product and Technology colleagues, contribute to technical interviews, and participate in the on-call rotation.
Requirements
- 6+ years of experience in a Site Reliability Engineer role.
- Strong knowledge of AWS and Azure cloud infrastructure, infrastructure as code, and orchestration tools.
- Experience implementing SLOs, SLIs, error budgets, observability, telemetry, centralized logging, and alerting.
- Production software development experience in any programming language.
- Experience with Terraform, Ansible, Docker, Kubernetes, CI/CD pipelines, artifact storage, GitHub, and cloud platforms.
- Experience with large-scale distributed systems and operational management of cloud software services.
Nice to have
- Associate or Professional certification from AWS or Microsoft Azure.
- Experience with Azure DevOps, Jenkins, GitLab, GitHub Actions, Artifactory, ECR, CloudWatch, Observe, or Prometheus.
Culture & Benefits
- Flexible working arrangements based on role and location.
- Employee stock options, retirement savings, support programs, and 401k matching.
- Health and dental plans, wellness sessions, an Employee Assistance Program, and Headspace subscription.
- Education subsidies, professional development programs, and learning opportunities.
- Generous leave options, including vacation, wellness, personal, parental, and volunteering time.
- Opportunities to explore responsible AI applications and collaborate with colleagues across global offices.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
10 дней назад
Site Reliability Engineer (AWS/Kubernetes)
90 000 - 120 000GBP
8 дней назад
Senior Site Reliability Engineer (Satellite Operations)
142 800 - 178 500$
6 дней назад
Senior Site Reliability Engineer (AWS)
7 дней назад
DevOps / SRE Engineer (Sheffield)
12 дней назад