Назад
Company hidden
24 минуты назад

Principal Engineer, CSRE Provisioning

Формат работы
remote (только United_kingdom)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal Engineer, CSRE Provisioning (SRE/platform infrastructure): Setting technical direction for operating, upgrading, migrating, and retiring Ticketmaster’s hosted CSRE platform estate with an accent on reliability, observability, disaster recovery, and infrastructure automation. Focus on leading high-risk migrations, resolving complex distributed-system failures, defining operating standards, and driving structural fixes for production incidents.

Location: Remote within the United Kingdom; the posting also references Farringdon, London.

Company

hirify.global operates Ticketmaster, a global live entertainment company and ticket marketplace connecting fans with live events.

What you will do

  • Set technical direction and a multi-quarter roadmap for lifecycle work across the CSRE hosted platform estate.
  • Lead the most complex and highest-risk platform initiatives, including upgrades, migrations, decommissioning, and service acceptance.
  • Define operating standards, supported patterns, and time-boxed exceptions for Provisioning.
  • Act as a technical escalation point for production incidents and turn incident analysis into durable resilience improvements.
  • Develop reliability, disaster recovery, observability, and controlled failure-testing practices across multiple platforms.
  • Represent Provisioning in architecture reviews, mentor senior and staff engineers, and participate in on-call rotations.

Requirements

  • Extensive experience operating and evolving production infrastructure at scale, including complex migrations and decommissioning.
  • Deep expertise in distributed systems, SRE principles, SLIs, SLOs, error budgets, and cross-service failure analysis.
  • Experience with on-premises data centers and cloud-native environments, including governance, cost trade-offs, and vendor evaluation.
  • Strong background in observability, alerting, infrastructure as code, automation, production readiness, and disaster recovery validation.
  • Excellent software engineering, technical communication, architecture documentation, and executive-summary skills.
  • Practical experience determining where LLM and AI-assisted tooling improves engineering work.

Culture & Benefits

  • Permanent, full-time employment with a 40-hour work week.
  • Work is guided by reliability, teamwork, integrity, and belonging.
  • Inclusive environment with support for balancing work and personal life.
  • Opportunities for professional development, learning, and career growth.
  • Hands-on ownership follows a “you build it, you run it” operating model.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →