Назад
Company hidden
обновлено 4 дня назад

Global Banking & Markets, Site Reliability Engineer, Vice President, Singapore (Cloud & AI)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Singapore
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Global Banking & Markets, Site Reliability Engineer, Vice President, Singapore (Cloud & AI): Building and operating highly available, multi-region cloud-native services for a 24×7 trading environment with an accent on reliability engineering, risk management, observability, and AI-assisted operations. Focus on designing event-driven trade lifecycle systems, coordinating complex incident response, and governing AI-generated automation for production use.

Location: Singapore, Singapore; office-based role

Company

hirify.global is a global investment banking, securities, and investment management firm serving clients, shareholders, and communities worldwide.

What you will do

  • Design, build, and operate high-availability, multi-region cloud-native services with security and comprehensive observability.
  • Define and manage SLIs, SLOs, error budgets, post-incident reviews, and durable reliability improvements.
  • Lead incident response for latency-sensitive, high-throughput trade lifecycle systems and coordinate global responders.
  • Develop event-driven architectures, processing pipelines, and optimized data paths for trading services.
  • Apply risk management to change management, capacity planning, resilience testing, failover, and business continuity drills.
  • Use AI coding and operations agents to accelerate diagnosis, automation, code comprehension, and delivery while maintaining production quality.

Requirements

  • 8+ years of professional software or reliability engineering experience and strong proficiency in at least one major programming language; Java 17+ is preferred.
  • Experience operating high-availability production environments with on-call duties, incident command, SLOs, error budgets, and post-incident reviews.
  • Strong understanding of GCP or AWS, Kubernetes, Docker, infrastructure as code, CI/CD, and automated delivery workflows.
  • Experience with AI models and AI-assisted engineering tools, including governing agents and critically evaluating generated output.
  • Experience building event-driven distributed systems with messaging platforms such as Apache Kafka.
  • Strong communication, stakeholder coordination, risk assessment, observability, and debugging skills across large codebases.

Nice to have

  • Experience with chaos engineering, capacity planning, load testing, and latency-sensitive production support.
  • Knowledge of Spring Boot, gRPC, Protocol Buffers, Apache Camel, Spring Integration, Terraform, Helm, Prometheus, Grafana, or OpenTelemetry.
  • Experience with data modeling, SQL/NoSQL databases, caching, JVM tuning, security patterns, mutual TLS, and secrets management.
  • Knowledge of equities, post-trade operations, trade lifecycle concepts, reconciliation, or financial services environments.

Culture & Benefits

  • Access to training and professional development opportunities.
  • Firmwide networks supporting professional and personal growth.
  • Benefits, wellness, personal finance, and mindfulness programs.
  • Commitment to diversity, inclusion, equal opportunity, and reasonable accommodations.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →