3 дня назад
Staff Site Reliability Engineer (SaaS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Site Reliability Engineer (SaaS): Building and operating reliable, scalable platform infrastructure for high-traffic commerce products with an accent on AWS, Infrastructure as Code, observability, and platform engineering. Focus on designing resilient systems, leading incident response, improving disaster recovery and cost efficiency, and enabling product teams to ship independently.
Location: Auckland, New Zealand; flexible hybrid workplace with remote work or office work available
Company
provides a cloud commerce platform that unifies online and physical operations, multichannel sales, payments, financial solutions, and supplier networks for retail, hospitality, and golf businesses.
What you will do
- Shape infrastructure reliability, developer experience, and platform capabilities.
- Design and maintain scalable, automated AWS infrastructure using Infrastructure as Code.
- Build platform tools and services that enable product engineers to ship and operate independently.
- Partner with engineering teams to design resilient, secure, and cost-effective systems.
- Improve observability, high availability, incident management, disaster recovery, and delivery practices.
- Lead incident response, participate in the on-call rotation, and mentor engineers.
Requirements
- Strong AWS experience operating highly available, scalable production systems.
- Experience in a SaaS or product-led environment and cross-functional collaboration.
- Strong Infrastructure as Code experience with Terraform and configuration management tools.
- Experience with Docker, Kubernetes, ECS, Linux, and complex shell scripting.
- Ability to code in Python, Ruby, or Go.
- Experience with observability, incident management, disaster recovery, security, datastores, cloud cost optimization, Agile, continuous delivery, and testing.
Culture & Benefits
- High autonomy and the opportunity to influence technical direction.
- Collaboration, knowledge sharing, and work across a broad range of infrastructure and reliability challenges.
- Flexible hybrid work with remote work available, plus the option to work remotely from anywhere in the world for up to 60 days per year.
- Share scheme, unlimited paid time off, health and wellness benefits, mental health support, and paid parental leave.
- Paid volunteer day, health insurance options, subsidized public transportation, and full-time parking.
- Newmarket office in Auckland with meals, snacks, events, and a dog-friendly environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
7 дней назад
Senior Site Reliability Engineer (Observability & Analytics) – Platform Infra
128 300 - 203 000CAD
3 дня назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
4 дня назад
Senior Site Reliability Engineer (AWS)
9 дней назад
Site Reliability Consultant (Kubernetes)
8 дней назад
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
95 000 - 145 000$
3 дня назад