Назад
Company hidden
2 часа назад

Senior Site Reliability Engineer (Fintech)

104 148 - 177 600$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Senior Site Reliability Engineer (Cloud/Python): Designing and maintaining the reliability, scalability, and observability of critical financial SaaS applications with an accent on SLOs/SLIs and automated incident response. Focus on building resilient cloud infrastructure on AWS/Azure and implementing AIOps to minimize toil and enable self-healing systems.

Location: Remote (US)

Salary: $104,148 – $177,600

Company

Financial SaaS provider focusing on secure and performant infrastructure for the fintech industry.

What you will do

  • Design and maintain SLOs and SLIs across critical systems to ensure consistent performance targets.
  • Lead observability strategy by designing comprehensive monitoring, logging, and tracing architectures.
  • Architect and deploy scalable cloud infrastructure on AWS and Azure using infrastructure-as-code practices.
  • Develop Python-based automation and AIOps capabilities to reduce toil and enable self-healing systems.
  • Drive reliability through load testing, chaos engineering, and failure scenario analysis.
  • Partner with backend teams on architecture reviews and implement defense-in-depth security for a regulated environment.

Requirements

  • Must be based in the US.
  • 7+ years of experience in SRE, DevOps, or platform engineering with production system responsibility.
  • Expert-level experience with AWS or Azure, including compute, networking, and storage.
  • Deep expertise in observability platforms such as Prometheus, Grafana, ELK, or Datadog.
  • Proficiency in Python and scripting languages (Bash, PowerShell) for production automation.
  • Proven track record of incident management and on-call ownership.

Nice to have

  • Experience in fintech, payments, or banking industries with knowledge of SOC 2 or PCI-DSS.
  • Proficiency with Kubernetes and container orchestration.
  • Background in chaos engineering using tools like Gremlin.
  • Contributions to open-source observability or infrastructure projects.

Culture & Benefits

  • Comprehensive insurance coverage (medical, dental, vision, life, and disability).
  • 401(k) plan with company match.
  • Flexible paid time off and paid holidays.
  • Fully remote work arrangement.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →