21 час назад
Senior Manager, Site Reliability Engineering (SRE) (Fintech)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Manager, Site Reliability Engineering (SRE) (Fintech): Leading SRE teams and defining reliability practices for a global, high-volume payments platform with an accent on availability, scalability, observability, and operational excellence. Focus on incident management, SLOs and error budgets, capacity planning, disaster recovery, chaos engineering, and automation for regulated financial infrastructure.
Location: Bangalore or Mumbai, India; hybrid, with 3 days per week in the office
Company
provides global infrastructure for real-time cross-border payments for banks, fintechs, and global businesses.
What you will do
- Lead, mentor, and grow SRE and reliability engineering teams across multiple time zones.
- Define and execute reliability strategy, including SLIs, SLOs, and error budgets for payment, card issuance, and compliance services.
- Own major incident management, on-call structures, escalation paths, and blameless postmortems.
- Build observability, capacity planning, performance engineering, automation, and self-healing systems for high-volume financial transactions.
- Own disaster recovery, business continuity, chaos engineering, production readiness reviews, runbooks, and operational standards.
- Partner with engineering, security, compliance, finance, and executive leadership on resilience, regulatory requirements, budgets, and risk reporting.
Requirements
- 10+ years of experience in software engineering, infrastructure, or SRE, including 4+ years leading SRE, DevOps, or infrastructure teams.
- Experience operating and scaling high-availability, transaction-heavy production systems; fintech, payments, banking, or e-commerce experience is preferred.
- Hands-on expertise with AWS, Kubernetes, container orchestration, and infrastructure as code such as Terraform or CloudFormation.
- Experience with observability stacks, alerting, SLOs, error budgets, distributed systems, databases, caching, messaging queues, and API-driven microservices.
- Experience leading major incident response and working with PCI-DSS, SOC 2, ISO 27001, and financial-services compliance requirements.
- Experience with Python, Go, or Bash automation and CI/CD tools such as Jenkins, ArgoCD, or GitLab CI; strong communication skills are required.
Nice to have
- Experience with real-time payment rails, card networks, or banking core integrations.
- Experience running SRE or infrastructure functions through hypergrowth or international expansion.
- FinOps and cloud cost optimization experience.
- Experience building an SRE function from the ground up.
Culture & Benefits
- Competitive salary, performance bonuses, equity for specific roles, and recognition programs.
- Medical coverage, a 24/7 employee assistance program, generous vacation, and a year-end shutdown.
- Hybrid working environment with three office days per week.
- Role-specific training, internal workshops, and a learning stipend.
- Social events, team bonding activities, happy hours, and team offsites.
- Global, diverse workplace with employees from more than 33 nationalities.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Senior Manager, Site Reliability Engineering (SRE)
Okta
9 часов назад
Manager, Site Reliability Engineering (Auth0) (Cloud Infrastructure)
182 000 - 250 800$
6 часов назад
Software Engineering Manager (Site Reliability Engineering)
110 110 - 204 490$
21 час назад
Sr Manager, Engineering - DevOps
7 дней назад
SVP, Head of Agentic Engineering and Acceleration (AI)
1 день назад