4 дня назад
Senior Site Reliability Engineer (AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AWS): Building and operating reliable, observable, and secure cloud infrastructure for a brand-tracking platform with an accent on golden paths, agentic systems, and operational readiness. Focus on designing least-privilege controls for agents and MCP servers, automating incident response and deployments, and balancing platform scalability, security, and cloud costs.
Location: Auckland, New Zealand; hybrid, office-first with flexibility for working from home and in-office
Company
provides always-on brand-tracking data and insights that help marketers measure brand growth, make decisions, and defend budgets.
What you will do
- Build and maintain AWS cloud infrastructure and paved paths with resilience, security, and scalability by default.
- Provide self-service monitoring, alerting, logging, tracing, provisioning, deployment, and operational automation.
- Improve incident response through tooling, runbooks, incident command, and post-incident reviews.
- Design identity and least-privilege controls for agents, CI systems, and MCP servers.
- Create infrastructure documentation and machine-readable operational runbooks for people and agents.
- Establish secure delivery guardrails, manage cloud and inference-cost visibility, and coach engineers on reliability practices.
Requirements
- Experience leading real production incidents and owning a platform through a meaningful scaling stage.
- Deep Infrastructure as Code experience with Terraform, Terragrunt, or CDK, plus containers, preferably ECS, cloud platforms, CI/CD, and Python, TypeScript, or Bash scripting.
- Strong observability skills with Datadog or similar tools, distributed tracing, structured logging, and incident management methodology.
- Knowledge of networking, cloud architecture, credential handling, least privilege, and prompt-injection risks in agentic systems.
- Experience using coding agents such as Claude Code and the ability to assess where they are effective.
- Product orientation, clear communication with technical and non-technical stakeholders, and a collaborative approach.
Nice to have
- Passion for marketing, design, and exceptional user experiences.
Culture & Benefits
- Transparent, trusting, supportive, and ambitious environment focused on learning and continuous improvement.
- Competitive market-rate compensation reviewed twice annually and an employee share option program.
- Wellness benefits, premium employee assistance support, and six weeks of paid annual leave.
- Paid parental leave, additional sick leave for IVF, and a gradual return-to-work option.
- $1,000 personal learning and development budget, with mentorship, speaking, and travel opportunities.
- Flexible working arrangements supported from offices in Auckland, Sydney, London, and New York.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 дней назад
Senior Site Reliability Engineer (AWS)
9 дней назад
Site Reliability Engineer (AWS/Terraform)
10 дней назад
Sr. Site Reliability Engineer
160 000 - 180 000$
8 дней назад
Principal Site Reliability Engineer (AI)
165 000 - 185 000$
7 дней назад
Senior Site Reliability Engineer (AI)
185 500 - 232 000$
9 дней назад