Назад
Company hidden
8 дней назад

Senior Site Reliability Engineer (AWS)

Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (AWS/Terraform): Building reliability standards, infrastructure automation, and observability tooling across on-premises private cloud and AWS environments with an accent on SLOs, production readiness, and incident response. Focus on leading complex reliability engineering, designing preventive automation, coordinating releases, and improving platform resilience across CCX, RCX, and HMS.

Location: 100% remote within the United States; candidates in the Mountain or Pacific Time zones are preferred. Travel is not required.

Company

hirify.global develops data quality, data enrichment, location intelligence, and customer communications management software.

What you will do

  • Define and maintain reliability standards, including SLOs, SLIs, error budgets, alerting, logging, and tracing.
  • Lead complex reliability engineering across on-premises private cloud and AWS environments supporting CCX, RCX, and HMS.
  • Build infrastructure-as-code, deployment automation, monitoring, alerting, and observability tooling with Terraform, Ansible, Datadog, and scripting languages.
  • Partner with engineering teams on scalability, operability, backup, recovery, failure modes, and production readiness.
  • Lead operational readiness reviews, release triage, deployment coordination, change reviews, and complex P1/P2 incident response.
  • Author root cause analyses, improve MTTR through preventive automation, maintain runbooks, and mentor SREs and engineering teams.

Requirements

  • Bachelor's degree in Computer Science, Information Systems, Engineering, or equivalent practical experience.
  • 5+ years of systems or infrastructure engineering experience in an enterprise production environment.
  • Advanced Linux proficiency across on-premises and cloud environments, with experience in complex infrastructure domains.
  • Proficiency with Terraform and Ansible role development, plus intermediate AWS experience with EC2, ECS, S3, VPC, IAM, CloudWatch, and Auto Scaling.
  • Proficiency in Python or Bash, with experience designing monitoring architecture, alerting strategies, CI/CD pipelines, and deployment automation.
  • Experience with TCP/IP, DNS, load balancing, distributed systems, SLOs, operational readiness reviews, root cause analysis, and cross-functional reliability planning.

Nice to have

  • Experience with Docker, Kubernetes, GitOps, Git, GitLab, enterprise virtualization, or ITIL practices.
  • Experience with Qualys, CrowdStrike, Rapid7, Wireshark, or protocol analysis.
  • AWS Solutions Architect, SysOps Administrator, or DevOps Engineer certification.
  • Prior experience mentoring engineers or leading on-call rotations.

Culture & Benefits

  • Fully remote work within the United States.
  • Rotating on-call coverage includes after-hours, weekends, and holidays within defined SLA windows.
  • Use hirify.global-provided AI tools such as GitHub Copilot or Claude for automation, incident analysis, runbook creation, solution testing, and architecture documentation.
  • Mentor engineers on AI-assisted engineering practices and operational standards.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →