24 минуты назад
Principal Engineer, CSRE Provisioning
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Engineer, CSRE Provisioning (SRE/platform infrastructure): Setting technical direction for operating, upgrading, migrating, and retiring Ticketmaster’s hosted CSRE platform estate with an accent on reliability, observability, disaster recovery, and infrastructure automation. Focus on leading high-risk migrations, resolving complex distributed-system failures, defining operating standards, and driving structural fixes for production incidents.
Location: Remote within the United Kingdom; the posting also references Farringdon, London.
Company
operates Ticketmaster, a global live entertainment company and ticket marketplace connecting fans with live events.
What you will do
- Set technical direction and a multi-quarter roadmap for lifecycle work across the CSRE hosted platform estate.
- Lead the most complex and highest-risk platform initiatives, including upgrades, migrations, decommissioning, and service acceptance.
- Define operating standards, supported patterns, and time-boxed exceptions for Provisioning.
- Act as a technical escalation point for production incidents and turn incident analysis into durable resilience improvements.
- Develop reliability, disaster recovery, observability, and controlled failure-testing practices across multiple platforms.
- Represent Provisioning in architecture reviews, mentor senior and staff engineers, and participate in on-call rotations.
Requirements
- Extensive experience operating and evolving production infrastructure at scale, including complex migrations and decommissioning.
- Deep expertise in distributed systems, SRE principles, SLIs, SLOs, error budgets, and cross-service failure analysis.
- Experience with on-premises data centers and cloud-native environments, including governance, cost trade-offs, and vendor evaluation.
- Strong background in observability, alerting, infrastructure as code, automation, production readiness, and disaster recovery validation.
- Excellent software engineering, technical communication, architecture documentation, and executive-summary skills.
- Practical experience determining where LLM and AI-assisted tooling improves engineering work.
Culture & Benefits
- Permanent, full-time employment with a 40-hour work week.
- Work is guided by reliability, teamwork, integrity, and belonging.
- Inclusive environment with support for balancing work and personal life.
- Opportunities for professional development, learning, and career growth.
- Hands-on ownership follows a “you build it, you run it” operating model.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →