8 часов назад
Staff Site Reliability Engineer (Fintech)
160 000 - 210 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Site Reliability Engineer (Kubernetes/Fintech): Ensuring the availability, performance, and scalability of revenue-generating trading systems with an accent on Kubernetes deployments, observability, infrastructure automation, and production reliability. Focus on leading SRE initiatives, designing SLIs/SLOs, automating incident response, and maintaining 99.95% uptime across web, mobile, desktop, and trading platforms.
Location: Chicago, IL; hybrid schedule with in-office work Tuesday through Thursday. Remote candidates are not accepted.
Salary: $160,000–$210,000 USD annually, plus a 12% annual target bonus.
Company
is a retail trading platform and futures broker serving more than 2 million users with trading software and brokerage services.
What you will do
- Lead the SRE function, set technical direction, and mentor engineers across reliability initiatives.
- Deploy and operate single- and multi-cluster services on Kubernetes.
- Analyze production incidents, perform root cause analysis, and implement remediation.
- Build automation for repetitive operational tasks, deployments, and incident response.
- Design monitoring and alerting systems and establish SLIs and SLOs across web, mobile, desktop, and trading platforms.
- Collaborate with engineering, operations, product, and QA teams to meet scalability, security, performance, and compliance requirements.
Requirements
- 8+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering.
- Expertise with Kubernetes, Docker, and container orchestration.
- Hands-on experience with CI/CD tools such as GitHub Actions.
- Proficiency in Python, Bash, or Go and automation tools such as Ansible, Terraform, or Helm.
- Cloud experience with AWS, GCP, or Azure, including networking, security, and identity management.
- Knowledge of monitoring and observability tools such as Prometheus, Grafana, or Datadog, plus strong leadership and communication skills.
Nice to have
- Trading industry experience.
- Contributions to open-source projects.
Culture & Benefits
- Weekly 12x7 on-call rotation, including weekend deployments and Sunday pre-market checkouts.
- Remote work on Mondays and Fridays, plus 20 additional flex remote days annually.
- Five company-wide office-optional weeks tied to major holidays.
- 23 days of annual paid time off, paid holidays, and one service day.
- 401(k) with a company match of up to 3.5%, health, vision, dental, life, and disability coverage, and paid parental bonding leave.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Senior Software Engineer, Site Reliability Engineering (AWS)
153 000 - 210 000$
SandboxAQ
5 дней назад
Staff Platform Engineer (AI)
121 600 - 228 000$
6 дней назад
Site Reliability Engineer (Healthcare)
85 000 - 130 000$
5 дней назад
Senior Site Reliability Engineer (Fintech)
129 000 - 175 000$
1 день назад
Senior Software Engineer (DevOps)
73 150 - 174 000$
4 дня назад
Site Reliability Engineer
87 400 - 123 400$