Назад
Company hidden
2 дня назад

Senior Site Reliability Engineer (Fintech)

180 000 - 200 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK/US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (Fintech): Building the reliability practice for a brokerage platform with an accent on customer-meaningful SLOs, error budgets, observability, and resilient Ruby, Java, and Elixir services. Focus on designing fault-injection exercises, scaling workloads across HashiCorp Nomad, and embedding reliability patterns into critical order execution and market data flows.

Location: Chicago, Illinois — hybrid, with 3 days per week in the office

Base salary: $180,000–$200,000 per year. Discretionary performance bonus: 15–20% of base salary.

Company

hirify.global is a Chicago-based retail brokerage and part of IG Group, a FTSE 100 fintech serving customers across five continents.

What you will do

  • Define customer-meaningful SLOs, error budgets, and multi-window burn-rate alerting for order execution and market data delivery.
  • Author reliability standards, including SLO methodology, error-budget policy, observability guidance, and Production Readiness Review checklists.
  • Implement reliability patterns such as circuit breakers, retries with backoff, bulkheads, and load shedding in Ruby, Java, and Elixir services.
  • Extend observability with Prometheus, Honeycomb, and OpenTelemetry while guiding workloads across the HashiCorp Nomad service fabric.
  • Design tabletop exercises and fault-injection tests for real-world failure and market-volatility scenarios.
  • Mentor engineers and build a cross-team culture of site reliability champions.

Requirements

  • Production-quality coding experience in Ruby and/or Java, plus Python for automation.
  • Experience embedding SRE practices within engineering teams, including SLOs, error budgets, and burn-rate alerting.
  • Hands-on experience with OpenTelemetry, Prometheus, and Grafana, including direct service instrumentation.
  • Strong Linux internals and networking fundamentals, including TCP/IP, UDP/multicast, packet capture, and flow analysis.
  • Production on-call experience and comfort establishing blameless post-incident review processes.
  • Experience influencing standards across teams without direct ownership.

Nice to have

  • Experience with HashiCorp Nomad, Consul, or Vault.

Culture & Benefits

  • Performance bonuses and stock purchase options.
  • Medical, vision, dental, 401(k), paid vacation, and paid sick leave.
  • Gym membership reimbursement, commuter benefits, pet insurance, and wellness and mental health programs.
  • Charitable donation matching and two paid volunteer days.
  • Daily catered lunch, office kitchen, in-building gym, and shuttle service to and from Metra.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →