7 дней назад
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra
128 300 - 203 000CAD
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra (Terraform/Python/Go): Operating and hardening shared Elastic Cloud observability infrastructure across hosted deployments with an accent on infrastructure as code, production reliability, and incident response. Focus on delivering complex platform projects, managing 24/7 on-call operations, improving Kubernetes-based systems, and building lasting fixes through RCAs and postmortems.
Location: Canada
Salary: $128,300–$203,000 CAD per year base salary, plus eligibility for stock participation and an RRSP match of up to 6% of eligible earnings.
Company
develops a cloud-based Search AI Platform covering search, security, and observability for organizations worldwide.
What you will do
- Own end-to-end delivery of moderate-to-high complexity projects with minimal day-to-day direction.
- Operate and harden shared Cloud infrastructure across ECH, ECE, and ECK using infrastructure as code.
- Write and review Terraform, Python, and Go used by platform engineers.
- Participate in a 24/7 on-call rotation, resolve incidents, and produce RCA and postmortem documentation.
- Review production designs and code, mentor less experienced engineers, and improve runbooks and operational processes.
Requirements
- 5+ years of experience in SRE, platform engineering, or infrastructure engineering.
- Strong Terraform proficiency, including ownership of large multi-workspace configurations.
- Strong Python software engineering fundamentals; experience with Go is beneficial.
- Deep Linux knowledge and experience operating containerized production workloads.
- Experience with 24/7 on-call rotations, incident resolution, and high-quality RCAs.
- Security-conscious infrastructure engineering, clear communication, cross-time-zone collaboration, and mentoring experience.
Nice to have
- Production experience with the Stack, including search, Logstash, Beats, and Kibana.
- Experience with ArgoCD, Helm, Kyverno, Vault, Teleport, Puppet, or Ansible.
- Exposure to FedRAMP, GovCloud, or other regulated infrastructure environments.
Culture & Benefits
- Flexible locations and schedules for eligible roles, with benefits designed for regional parity.
- Health coverage for employees and families in many locations.
- Generous vacation allowance and at least 16 weeks of parental leave.
- Up to $2,000 in matched financial donations and up to 40 volunteer hours annually.
- Security and privacy responsibilities aligned with 's Secure Software Development Framework.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
95 000 - 145 000$
9 дней назад
Site Reliability Consultant (Kubernetes)
3 дня назад
Zscaler Site Reliability Engineer (SRE)
10 дней назад
Senior Site Reliability Engineer (Azure)
99 840 - 137 280CAD
9 дней назад
SRE Monitoring Platform Software Engineer (AI)
14 дней назад
Service Reliability Engineer (Kubernetes)
180 000 - 210 000$