Назад
Company hidden
9 дней назад

Site Reliability Engineering (SRE) Manager

123 650 - 229 650$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior/lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineering (SRE) Manager (AWS/Kubernetes): Leading North American SRE teams and owning reliability strategy across AWS, GCP, EKS, and GKE platforms with an accent on incident management, automation, infrastructure as code, and observability. Focus on designing self-healing systems, coordinating high-severity incidents, improving ITSM processes, and scaling monitoring and operational readiness across multiple regions.

Location: Onsite in Frisco, Texas; candidates must live within a commutable distance of the Frisco office.

Salary: USD $123,650–$229,650 per year, depending on experience and qualifications.

Company

hirify.global develops consumer security solutions designed to protect people, families, and communities in an always-online world.

What you will do

  • Lead, hire, mentor, and develop the North American SRE team while setting technical direction and reliability strategy.
  • Own Incident and Problem Management, including executive communication, post-incident reviews, escalation paths, and on-call readiness.
  • Guide reliability engineering across AWS, GCP, EKS, and GKE platforms, including troubleshooting complex Kubernetes operational issues.
  • Drive Python automation, self-healing tooling, runbooks, and Terraform-based infrastructure provisioning to reduce manual toil.
  • Define observability standards using Grafana, alerting, SQL, and CloudWatch Logs Insights across multiple deployment regions.
  • Act as incident commander during critical events and report reliability metrics, risks, trends, and remediation roadmaps to senior leadership.

Requirements

  • 9+ years of experience in SRE, DevOps, infrastructure, or related roles, including significant leadership or management experience.
  • Experience building and leading high-performing technical teams and managing hiring, performance, and career development.
  • Hands-on AWS experience and a strong background troubleshooting EKS/GKE and Kubernetes environments.
  • Technical fluency in Python, Terraform, Grafana, SQL, CloudWatch Logs Insights, monitoring, and alerting.
  • Organizational ownership of ITSM, Incident Management, and Problem Management processes, with strong executive and stakeholder communication skills.
  • AWS Certification and Certified Kubernetes Administrator (CKA) or equivalent Kubernetes certification are required.

Nice to have

  • Bachelor’s degree in Computer Science, Information Technology, or a related field.
  • Master’s degree or MBA.

Culture & Benefits

  • Flexible work hours and family-friendly benefits.
  • Bonus program and 401(k) retirement plan.
  • Medical, dental, vision, life, and disability coverage.
  • Paid parental leave, paid company holidays, sick time, and vacation.
  • Social programs, community involvement, and a focus on diversity and inclusion.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →