обновлено 10 дней назад
Site Reliability Engineer, Global Banking & Markets, Vice President
150 000 - 250 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer, Global Banking & Markets, Vice President (Cloud/SRE): Building and operating high-availability, multi-region cloud-native services supporting latency-sensitive trading systems with an accent on reliability engineering, observability, distributed architectures, and AI-assisted operations. Focus on defining SLOs and error budgets, leading incident response, designing event-driven processing pipelines, and strengthening resilience through risk management, chaos engineering, and failover testing.
Location: New York, New York, United States
Salary: $150,000–$250,000 USD base salary per year, plus potential discretionary bonus.
Company
is a global investment banking, securities, and investment management firm operating across financial markets.
What you will do
- Design, build, and operate high-availability, multi-region cloud-native services for global 24×7 trading operations.
- Define and manage SLIs, SLOs, error budgets, reliability standards, observability, and post-incident improvements.
- Lead incident response for latency-sensitive, high-throughput trade lifecycle systems and coordinate technical and business stakeholders.
- Develop event-driven architectures, multi-stage processing pipelines, optimized data paths, and resilient distributed systems.
- Apply risk management to change control, capacity planning, resilience testing, chaos engineering, failover, and business continuity drills.
- Use AI-assisted engineering and operations agents to accelerate root-cause analysis, code comprehension, remediation, and automation while governing output quality.
Requirements
- 8+ years of professional software or reliability engineering experience.
- Strong command of at least one major programming language; Java 17+ is preferred.
- Experience operating high-availability production environments, including on-call, incident command, SLOs, error budgets, and post-incident reviews.
- Strong knowledge of cloud infrastructure, including GCP or AWS, Kubernetes, Docker, and infrastructure-as-code.
- Experience with AI models and AI-assisted engineering tools, including critical evaluation and governance of AI-generated work.
- Experience with event-driven and distributed systems, CI/CD automation, observability, messaging platforms such as Apache Kafka, and production debugging.
Nice to have
- Chaos engineering, capacity planning, load testing, and latency-sensitive production support.
- Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, or OpenTelemetry.
- AI applications for anomaly detection, log analysis, automated remediation, and agentic operations.
- SQL or NoSQL databases, caching, performance optimization, enterprise security, mutual TLS, secrets management, or certificate rotation.
- Equities, post-trade, trade lifecycle, position management, reconciliation, or financial services experience.
Culture & Benefits
- Work within a global engineering organization supporting financial markets and distributed stakeholders.
- Benefits and wellness offerings are generally available to eligible active US employees working at least 20 hours per week.
- Training, professional development, firmwide networks, wellness programs, and personal finance offerings are available.
- Reasonable accommodations are available during the recruiting process for candidates with disabilities or special needs.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
14 дней назад
Staff Software Engineer, Reliability
203 500 - 248 500$
12 дней назад
Senior Site Reliability Engineer (AI Platform)
69 768 - 96 900€
14 дней назад
Director, Site Reliability (Fintech)
244 375 - 316 250$
13 дней назад
Staff Site Reliability Engineer (AI Ops)
152 000 - 228 000$
10 дней назад
Senior Software Engineer, Site Reliability
180 000 - 225 000$
11 дней назад
Site Reliability Engineer, Observability
160 000 - 200 000$