6 дней назад
Vice President - Site Reliability Engineering (SRE) – The Core Engineering
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Vice President - Site Reliability Engineering (SRE) – The Core Engineering (SRE/Cloud Infrastructure): Engineering highly reliable, observable, and resilient platforms for critical business services with an accent on distributed systems, automation, and production excellence. Focus on designing fault-tolerant architectures, leading complex incidents, reducing operational toil, and establishing SLOs, SLIs, and error budgets.
Location: New York, NY, United States
Company
is a global investment banking, securities, and investment management firm providing financial services to corporations, financial institutions, governments, and individuals.
What you will do
- Define SLOs, SLIs, and error budgets with engineering leadership.
- Architect highly available, fault-tolerant, and self-healing systems with product development teams.
- Build automation, tooling, and self-service capabilities to reduce operational toil.
- Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.
- Lead complex, multi-system production incidents and facilitate blameless post-mortems.
- Design sustainable on-call models, escalation paths, and balanced pager responsibilities.
Requirements
- Strong proficiency in Java, Python, Node.js, or another major programming language.
- Hands-on experience with Infrastructure as Code, including Terraform, Ansible, or CloudFormation.
- Deep knowledge of Docker, Kubernetes, service meshes, and ingress controllers.
- Advanced experience with AWS, GCP, or Azure cloud-native architectures.
- Experience with observability, automated testing, SDLC concepts, Linux, networking, load balancing, algorithms, data structures, and software design.
- 7–10 years of experience and a bachelor's degree in Computer Science, Systems Engineering, or a related technical field are preferred.
Nice to have
- Interest in financial markets and technology.
- Experience influencing architecture and driving SRE adoption without direct authority.
Culture & Benefits
- Blameless culture focused on learning from failures and improving systems.
- Fast-paced environment with collaboration across multiple engineering teams.
- Emphasis on sustainable operations, clear escalation paths, and healthy on-call practices.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
7 дней назад
Sr Software Development Engineer, SRE (US Federal)
163 800 - 245 800$
8 дней назад
Senior DevOps / Site Reliability Engineer (SRE) (Cybersecurity)
165 000 - 215 000$
12 дней назад
Staff Site Reliability Engineer (AI)
252 000 - 308 000$
11 дней назад
Principal SRE (Site Reliability Engineering)
137 400 - 233 600$
Phantom
9 дней назад
Staff Software Engineer (SRE) (Crypto)
200 000 - 250 000$
6 дней назад