21 Π΄Π΅Π½Ρ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (PostgreSQL)
ΠΡΡΡ & Π‘ΠΎΠΏΡΠΎΠ²ΠΎΠ΄
ΠΠ»Ρ ΠΌΡΡΡΠ° Ρ ΡΡΠΎΠΉ Π²Π°ΠΊΠ°Π½ΡΠΈΠ΅ΠΉ Π½ΡΠΆΠ΅Π½ Plus
ΠΠΏΠΈΡΠ°Π½ΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
Π’Π΅ΠΊΡΡ:
TL;DR
Senior Site Reliability Engineer (PostgreSQL): Operating and improving reliable, observable brokerage infrastructure across cloud services, Kubernetes, messaging, and data platforms with an accent on PostgreSQL reliability, production operations, and infrastructure automation. Focus on designing SLOs, resolving complex incidents, performing safe online migrations, and strengthening high-availability and disaster-recovery capabilities for trading-critical systems.
Location: Remote - Americas
Company
provides agent-first brokerage infrastructure and developer-friendly APIs for stocks, ETFs, options, crypto, fixed income, and other financial services.
What you will do
- Operate production systems through on-call coverage, incident response, postmortems, and remediation follow-ups.
- Define and refine SLIs, SLOs, and error budgets with product engineering teams.
- Improve observability across metrics, logs, traces, and alerting.
- Ship cloud resources and Kubernetes workloads through infrastructure as code and GitOps workflows.
- Own PostgreSQL reliability work, including performance tuning, schema and migration reviews, online migrations, HA/DR, and CDC pipelines.
- Mentor engineers through code reviews, design reviews, and pairing on reliability and database practices.
Requirements
- 4+ years of experience in SRE, DevOps, platform, infrastructure, or backend engineering with significant production operations ownership.
- Hands-on experience operating production services on Kubernetes and delivering infrastructure through GitOps.
- Production PostgreSQL expertise, including query plans, pg_stat_*, indexing, schema trade-offs, and safe online migrations.
- Knowledge of cloud networking fundamentals, including VPCs, routing, L4/L7 load balancing, DNS, and TLS.
- Experience with observability tooling, Linux operations, and structured incident response.
- Working proficiency in Go or Python, strong communication skills, and an interest in developing deeper PostgreSQL and DBA expertise.
Nice to have
- Experience with large PostgreSQL clusters, OLTP workloads, online migrations, HA/DR, connection pooling, or CDC pipelines.
- Experience with typed SQL access layers in Go, such as pgx, gorm, or sqlc.
- Production experience with RabbitMQ, Kafka, or Redpanda.
- Security and compliance experience in regulated environments.
- Experience in trading, brokerage, or regulated fintech.
Culture & Benefits
- Work with a globally distributed team of 400+ professionals.
- Competitive salary and stock options.
- Health benefits.
- One-time USD $500 home-office setup allowance for new hires.
- USD $150 monthly stipend through a Brex Card.
ΠΡΠ΄ΡΡΠ΅ ΠΎΡΡΠΎΡΠΎΠΆΠ½Ρ: Π΅ΡΠ»ΠΈ ΡΠ°Π±ΠΎΡΠΎΠ΄Π°ΡΠ΅Π»Ρ ΠΏΡΠΎΡΠΈΡ Π²ΠΎΠΉΡΠΈ Π² ΠΈΡ ΡΠΈΡΡΠ΅ΠΌΡ, ΠΈΡΠΏΠΎΠ»ΡΠ·ΡΡ iCloud/Google, ΠΏΡΠΈΡΠ»Π°ΡΡ ΠΊΠΎΠ΄/ΠΏΠ°ΡΠΎΠ»Ρ, Π·Π°ΠΏΡΡΡΠΈΡΡ ΠΊΠΎΠ΄/ΠΠ, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡΠ΅ ΡΡΠΎΠ³ΠΎ - ΡΡΠΎ ΠΌΠΎΡΠ΅Π½Π½ΠΈΠΊΠΈ. ΠΠ±ΡΠ·Π°ΡΠ΅Π»ΡΠ½ΠΎ ΠΆΠΌΠΈΡΠ΅ "ΠΠΎΠΆΠ°Π»ΠΎΠ²Π°ΡΡΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡΠΈΡΠ΅ Π² ΠΏΠΎΠ΄Π΄Π΅ΡΠΆΠΊΡ. ΠΠΎΠ΄ΡΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β
ΠΠΎΡ ΠΎΠΆΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Staff Site Reliability Engineer (AWS GovCloud)
158Β 500 - 230Β 000$
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer, GovCloud (AWS)
128Β 500 - 190Β 000$
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Senior Software Engineer, Cloud Infrastructure (AWS)
180Β 000 - 225Β 000$
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (AI)
4 Π΄Π½Ρ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (Kubernetes)
6 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄