Назад
Company hidden
4 дня назад

Senior Site Reliability Engineer (AWS)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK/US/Australia +1 еще
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (AWS): Building and operating reliable, observable, and secure cloud infrastructure for a brand-tracking platform with an accent on golden paths, agentic systems, and operational readiness. Focus on designing least-privilege controls for agents and MCP servers, automating incident response and deployments, and balancing platform scalability, security, and cloud costs.

Location: Auckland, New Zealand; hybrid, office-first with flexibility for working from home and in-office

Company

hirify.global provides always-on brand-tracking data and insights that help marketers measure brand growth, make decisions, and defend budgets.

What you will do

  • Build and maintain AWS cloud infrastructure and paved paths with resilience, security, and scalability by default.
  • Provide self-service monitoring, alerting, logging, tracing, provisioning, deployment, and operational automation.
  • Improve incident response through tooling, runbooks, incident command, and post-incident reviews.
  • Design identity and least-privilege controls for agents, CI systems, and MCP servers.
  • Create infrastructure documentation and machine-readable operational runbooks for people and agents.
  • Establish secure delivery guardrails, manage cloud and inference-cost visibility, and coach engineers on reliability practices.

Requirements

  • Experience leading real production incidents and owning a platform through a meaningful scaling stage.
  • Deep Infrastructure as Code experience with Terraform, Terragrunt, or CDK, plus containers, preferably ECS, cloud platforms, CI/CD, and Python, TypeScript, or Bash scripting.
  • Strong observability skills with Datadog or similar tools, distributed tracing, structured logging, and incident management methodology.
  • Knowledge of networking, cloud architecture, credential handling, least privilege, and prompt-injection risks in agentic systems.
  • Experience using coding agents such as Claude Code and the ability to assess where they are effective.
  • Product orientation, clear communication with technical and non-technical stakeholders, and a collaborative approach.

Nice to have

  • Passion for marketing, design, and exceptional user experiences.

Culture & Benefits

  • Transparent, trusting, supportive, and ambitious environment focused on learning and continuous improvement.
  • Competitive market-rate compensation reviewed twice annually and an employee share option program.
  • Wellness benefits, premium employee assistance support, and six weeks of paid annual leave.
  • Paid parental leave, additional sick leave for IVF, and a gradual return-to-work option.
  • $1,000 personal learning and development budget, with mentorship, speaking, and travel opportunities.
  • Flexible working arrangements supported from offices in Auckland, Sydney, London, and New York.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →