Назад
22 часа назад

Manager, Software Engineering - SRE Site Lead: Dublin ROC (AI)

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
Ireland
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Manager, Software Engineering - SRE Site Lead: Dublin ROC (AI): Leading a Dublin site that provides 24/7 live incident response, software triage, and incident mitigation for Riot games with an accent on SRE practices, high-availability systems, and engineering team development. Focus on building incident response automation, reducing alert fatigue, coordinating critical incident triage, and developing an AI-focused approach to systems triage.

Location: Dublin, Ireland

Company

Riot Games develops games and player experiences, including globally operated live services.

What you will do

  • Manage the Dublin ROC engineering staff, including hiring, performance management, mentorship, and career development.
  • Lead 24/7 follow-the-sun incident response coverage and participate in the on-call rotation as an Incident Commander.
  • Act as the technical lead through hands-on engineering, code reviews, design reviews, and incident response automation.
  • Manage project and incident-response capacity while tracking and reporting site performance metrics.
  • Drive SRE best practices, reduce alert fatigue, and improve triage and early problem detection.
  • Lead stakeholder coordination during critical incidents and major launches, working with TPM teams.

Requirements

  • Bachelor’s or Master’s degree in Computer Science or a related field, or equivalent professional experience.
  • At least 2 years of experience as a Senior Software Engineer or above, plus at least 2 years managing engineers through hiring, coaching, and career development.
  • Experience triaging large production systems, acting as an Incident Commander, and eliminating alert fatigue through effective alerting design.
  • Experience leading SRE practices and designing, prioritizing, and maintaining high-capacity, high-availability, high-performance backend services.
  • Experience with container-based ecosystems and schedulers such as Kubernetes, GKE, Amazon ECS, Marathon, or Mesos.
  • Ability to work across multiple organizations and align teams on technical standards.

Nice to have

  • At least 4 years of experience in a high-performance Site Reliability capacity.
  • Experience with global follow-the-sun operations and communication across time zones.
  • Experience with distributed systems, microservices, relational databases such as MySQL, CI/CD pipelines, and AWS or comparable cloud environments.
  • Understanding of software performance and latency in online games.

Culture & Benefits

  • Collaborative environment focused on player experience and player empathy.
  • Open paid time off policy and flexible work schedules.
  • Medical, dental, and life insurance.
  • Parental leave for employees, spouses or domestic partners, and children.
  • Retirement benefits with a company match and charitable donation support.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →