обновлено 4 дня назад
Global Banking & Markets, Site Reliability Engineer, Vice President, Singapore (Cloud & AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Global Banking & Markets, Site Reliability Engineer, Vice President, Singapore (Cloud & AI): Building and operating highly available, multi-region cloud-native services for a 24×7 trading environment with an accent on reliability engineering, risk management, observability, and AI-assisted operations. Focus on designing event-driven trade lifecycle systems, coordinating complex incident response, and governing AI-generated automation for production use.
Location: Singapore, Singapore; office-based role
Company
is a global investment banking, securities, and investment management firm serving clients, shareholders, and communities worldwide.
What you will do
- Design, build, and operate high-availability, multi-region cloud-native services with security and comprehensive observability.
- Define and manage SLIs, SLOs, error budgets, post-incident reviews, and durable reliability improvements.
- Lead incident response for latency-sensitive, high-throughput trade lifecycle systems and coordinate global responders.
- Develop event-driven architectures, processing pipelines, and optimized data paths for trading services.
- Apply risk management to change management, capacity planning, resilience testing, failover, and business continuity drills.
- Use AI coding and operations agents to accelerate diagnosis, automation, code comprehension, and delivery while maintaining production quality.
Requirements
- 8+ years of professional software or reliability engineering experience and strong proficiency in at least one major programming language; Java 17+ is preferred.
- Experience operating high-availability production environments with on-call duties, incident command, SLOs, error budgets, and post-incident reviews.
- Strong understanding of GCP or AWS, Kubernetes, Docker, infrastructure as code, CI/CD, and automated delivery workflows.
- Experience with AI models and AI-assisted engineering tools, including governing agents and critically evaluating generated output.
- Experience building event-driven distributed systems with messaging platforms such as Apache Kafka.
- Strong communication, stakeholder coordination, risk assessment, observability, and debugging skills across large codebases.
Nice to have
- Experience with chaos engineering, capacity planning, load testing, and latency-sensitive production support.
- Knowledge of Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, or OpenTelemetry.
- Experience with data modeling, SQL/NoSQL databases, caching, JVM tuning, security patterns, mutual TLS, and secrets management.
- Knowledge of equities, post-trade operations, trade lifecycle concepts, reconciliation, or financial services environments.
Culture & Benefits
- Access to training and professional development opportunities.
- Firmwide networks supporting professional and personal growth.
- Benefits, wellness, personal finance, and mindfulness programs.
- Commitment to diversity, inclusion, equal opportunity, and reasonable accommodations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 часов назад
Cloud Infra Engineer
9 часов назад
System Reliability Engineer (Data Centre, Cloud Infrastructure)
8 часов назад
Senior / Staff Platform Engineer (AI)
10 часов назад
Software Production Engineer (Fintech)
10 часов назад
Infra Platform Engineer (Cloud Infrastructure)
2 часа назад