Назад
Company hidden
5 дней назад

Senior Production Engineer (Fintech)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Production Engineer (Fintech): Building and operating reliable production systems, observability platforms, automation, and self-healing capabilities for a cloud-native capital markets platform with an accent on incident recovery, infrastructure resilience, and operational tooling. Focus on reducing MTTD and MTTR, designing monitoring and diagnostic systems, and applying SRE, Infrastructure as Code, and GitOps practices to eliminate operational toil.

Location: London, UK

Company

hirify.global builds a cloud-native, end-to-end capital markets platform that provides sophisticated investors with access to assets across global markets.

What you will do

  • Own the health, resilience, and recovery of production systems while responding to incidents and improving operational reliability.
  • Design monitoring and observability platforms that reduce alert fatigue and accelerate root-cause analysis.
  • Develop automation, self-healing workflows, diagnostic tools, runbooks, and recovery playbooks that reduce manual toil and MTTR.
  • Analyze incidents and systemic failure patterns, then build solutions that prevent recurring classes of failures.
  • Partner with Platform Engineering on CI/CD pipelines, deployment safety, Infrastructure as Code, GitOps, and infrastructure resilience.
  • Measure production health through SLIs, SLOs, and SLAs, and explore AI-assisted diagnostics and developer tooling.

Requirements

  • Strong hands-on Python skills for automation and tooling.
  • Experience in SRE, Production Engineering, Platform Engineering, or a related discipline with direct production ownership.
  • Experience building automation and diagnostic tooling that improves recovery times or reduces operational toil.
  • Deep familiarity with cloud-native technologies, including Kubernetes, containers, and distributed systems.
  • Experience with observability platforms such as Datadog, plus exposure to Terraform and GitOps workflows using ArgoCD, GitHub Actions, or similar tools.
  • Ability to solve ambiguous, high-stakes production problems and communicate effectively across engineering, operations, and business teams.

Nice to have

  • Familiarity with Java, Go, Kafka, Redis, Snowflake, PostgreSQL, gRPC, Protobuf, and internal platform APIs.
  • Fintech or financial industry experience.

Culture & Benefits

  • Reliability is treated as a product, with uptime and engineer experience as core measures.
  • Production Engineering combines SRE discipline with a builder mindset focused on lasting solutions rather than workarounds.
  • Work spans engineering and operations teams, with measurable impact on platform reliability and recovery speed.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →