22 часа назад
Manager, Software Engineering - SRE Site Lead: Dublin ROC (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Manager, Software Engineering - SRE Site Lead: Dublin ROC (AI): Leading a Dublin site that provides 24/7 live incident response, software triage, and incident mitigation for Riot games with an accent on SRE practices, high-availability systems, and engineering team development. Focus on building incident response automation, reducing alert fatigue, coordinating critical incident triage, and developing an AI-focused approach to systems triage.
Location: Dublin, Ireland
Company
Riot Games develops games and player experiences, including globally operated live services.
What you will do
- Manage the Dublin ROC engineering staff, including hiring, performance management, mentorship, and career development.
- Lead 24/7 follow-the-sun incident response coverage and participate in the on-call rotation as an Incident Commander.
- Act as the technical lead through hands-on engineering, code reviews, design reviews, and incident response automation.
- Manage project and incident-response capacity while tracking and reporting site performance metrics.
- Drive SRE best practices, reduce alert fatigue, and improve triage and early problem detection.
- Lead stakeholder coordination during critical incidents and major launches, working with TPM teams.
Requirements
- Bachelor’s or Master’s degree in Computer Science or a related field, or equivalent professional experience.
- At least 2 years of experience as a Senior Software Engineer or above, plus at least 2 years managing engineers through hiring, coaching, and career development.
- Experience triaging large production systems, acting as an Incident Commander, and eliminating alert fatigue through effective alerting design.
- Experience leading SRE practices and designing, prioritizing, and maintaining high-capacity, high-availability, high-performance backend services.
- Experience with container-based ecosystems and schedulers such as Kubernetes, GKE, Amazon ECS, Marathon, or Mesos.
- Ability to work across multiple organizations and align teams on technical standards.
Nice to have
- At least 4 years of experience in a high-performance Site Reliability capacity.
- Experience with global follow-the-sun operations and communication across time zones.
- Experience with distributed systems, microservices, relational databases such as MySQL, CI/CD pipelines, and AWS or comparable cloud environments.
- Understanding of software performance and latency in online games.
Culture & Benefits
- Collaborative environment focused on player experience and player empathy.
- Open paid time off policy and flexible work schedules.
- Medical, dental, and life insurance.
- Parental leave for employees, spouses or domestic partners, and children.
- Retirement benefits with a company match and charitable donation support.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
Staff Site Reliability Engineer (Cloud Infrastructure)
2 дня назад
Senior Platform Engineering Manager (AI)
183 000 - 229 000$
7 дней назад
Staff Site Reliability Engineer
7 дней назад
Software Engineering Manager, Backend (.NET)
6 дней назад
Sr. Software Development Manager, Relational Database Services (AI)
2 дня назад
Infrastructure Engineering Director (iGaming)
197 000 - 246 000$