4 дня назад
Staff Site Reliability Expert (AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Site Reliability Expert (AWS): Building and operating scalable AWS infrastructure and platform services for high-traffic commerce products with an accent on reliability, automation, observability, and developer experience. Focus on designing resilient systems, leading incident response, improving disaster recovery and security, and driving platform-wide engineering practices.
Location: Auckland, New Zealand; hybrid workplace with remote work options
Company
provides a cloud commerce platform that unifies online and physical operations, multichannel sales, payments, financial solutions, and supplier networks for retail, hospitality, and golf businesses.
What you will do
- Shape infrastructure reliability, developer experience, and platform capabilities.
- Design and maintain scalable, automated AWS infrastructure with Infrastructure as Code.
- Build platform services and tools that help product engineers ship and operate independently.
- Partner with engineering teams to design resilient, secure, and cost-effective systems.
- Champion observability, high availability, incident management, disaster recovery, and cloud cost optimisation.
- Lead incident response, drive lasting improvements, mentor engineers, and participate in the on-call rotation.
Requirements
- Strong experience operating highly available and scalable production systems on AWS.
- Experience in a SaaS or product-led environment and cross-functional collaboration.
- Strong Infrastructure as Code experience with Terraform and configuration management tooling.
- Experience with Docker, Kubernetes, ECS, Linux, and complex shell scripting.
- Ability to code in Python, Ruby, or Go.
- Experience with observability, incident management, disaster recovery, security, datastores, continuous delivery, testing, and cloud cost optimisation.
Culture & Benefits
- High autonomy and the opportunity to influence technical direction across teams.
- Flexible hybrid workplace with remote work options.
- Remote work from anywhere in the world for up to 60 days per year.
- Share scheme, unlimited paid time off, health and wellness benefit, and mental health support.
- Paid parental leave, volunteer day, health insurance options, and subsidised public transportation.
- Modern dog-friendly office in Newmarket, Auckland, with meals, snacks, events, and full-time parking.
Hiring process
- Applicants should disclose criminal convictions; criminal record checks are part of the hiring process.
- Workplace accommodations are available on request.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Site Reliability Engineer (Azure)
99 840 - 137 280CAD
5 дней назад
Senior Site Reliability Engineer (Satellite Operations)
142 800 - 178 500$
6 дней назад
Senior Operations Reliability Engineer (IAM)
92 000 - 120 800$
10 дней назад
Sr. Analyst, Application Reliability Engineering - Operations
Chainlink Labs
4 дня назад
Senior Site Reliability Engineer, DevEx (Kubernetes)
90 000 - 125 000$
Okta
8 дней назад
Senior Software Reliability Engineer - Data Platform
136 000 - 187 000CAD