2 дня назад
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Vice President - Site Reliability Engineering (SRE) (Cloud-Native Platforms): Engineering highly reliable, observable, and resilient platforms for critical business services with an accent on distributed systems, automation, and production excellence. Focus on designing fault-tolerant architectures, leading complex multi-system incident response, and improving reliability through chaos engineering, performance tuning, and capacity forecasting.
Location: Birmingham, West Midlands, England, United Kingdom
Company
is a global investment banking, securities, and investment management firm providing financial services to corporations, financial institutions, governments, and individuals.
What you will do
- Establish service level objectives, service level indicators, and error budgets with engineering leadership.
- Architect highly available, fault-tolerant, and self-healing systems with product development teams.
- Build automation, tooling, and self-service capabilities to reduce operational toil.
- Improve production readiness through load testing, performance tuning, capacity forecasting, chaos engineering, and reliability reviews.
- Lead complex, multi-system production incidents and facilitate blameless post-mortems.
- Design sustainable on-call models, escalation paths, and balanced pager responsibilities.
Requirements
- Strong proficiency in at least one major programming language, such as Java, Python, or Node.
- Hands-on experience with Infrastructure as Code tools including Terraform, Ansible, or CloudFormation.
- Deep knowledge of Docker, Kubernetes, service meshes, and ingress controllers.
- Advanced experience with AWS, GCP, or Azure and resilient cloud-native architectures.
- Experience with observability technologies for distributed tracing, logging, and metrics, plus automated testing and Linux-based development.
- 7–10 years of experience and preferably a bachelor's degree in Computer Science, Systems Engineering, or a related technical field.
Nice to have
- Knowledge of networking protocols, VPCs, load balancing, database query performance tuning, algorithms, data structures, and software design.
- Interest in financial markets and technology.
Culture & Benefits
- Blameless culture focused on learning from failures and improving systems.
- Opportunities for professional and personal development through training and firmwide networks.
- Benefits, wellness, personal finance, and mindfulness programs.
- Commitment to diversity, inclusion, and reasonable accommodations during recruitment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Sr. Site Reliability Engineer (Healthcare)
125 000 - 145 000$
6 дней назад
Senior Site Reliability Engineer (SRE Infrastructure)
5 дней назад
Sr Principal Site Reliability Engineer (Sovereign Cloud)
1 день назад
Site Reliability Engineer (AI SaaS)
4 дня назад
Staff Site Reliability Engineer (AI/Blockchain)
195 000 - 257 500$
3 дня назад
Senior Site Reliability Engineer (Fintech)
160 000 - 200 000$