Назад
Company hidden
2 дня назад

Site Reliability Engineer III (GCP)

Π€ΠΎΡ€ΠΌΠ°Ρ‚ Ρ€Π°Π±ΠΎΡ‚Ρ‹
hybrid
Π’ΠΈΠΏ Ρ€Π°Π±ΠΎΡ‚Ρ‹
fulltime
Π“Ρ€Π΅ΠΉΠ΄
senior
Английский
b2
Π‘Ρ‚Ρ€Π°Π½Π°
UK
Вакансия ΠΈΠ· списка Hirify.GlobalВакансия ΠΈΠ· Hirify Global, списка ΠΌΠ΅ΠΆΠ΄ΡƒΠ½Π°Ρ€ΠΎΠ΄Π½Ρ‹Ρ… tech-ΠΊΠΎΠΌΠΏΠ°Π½ΠΈΠΉ
Для мэтча ΠΈ ΠΎΡ‚ΠΊΠ»ΠΈΠΊΠ° Π½ΡƒΠΆΠ΅Π½ Plus

ΠœΡΡ‚Ρ‡ & Π‘ΠΎΠΏΡ€ΠΎΠ²ΠΎΠ΄

Для мэтча с этой вакансиСй Π½ΡƒΠΆΠ΅Π½ Plus

ОписаниС вакансии

ВСкст:
/
TL;DR
Site Reliability Engineer III (GCP): Building and operating resilient, automated GCP infrastructure and middleware platforms for clearing, risk, and derivatives applications with an accent on ultra-low-latency performance, high concurrency, and observability. Focus on designing reliability tooling, leading production incident recovery, reducing operational toil through automation, and testing disaster recovery and system resiliency.

Location: Belfast, United Kingdom; hybrid working

Company

hirify.global operates a leading global derivatives marketplace and builds technology for clearing, risk, and derivatives applications.

What you will do

  • Architect, operate, and migrate messaging, service discovery, and data distribution platforms to Google Cloud.
  • Manage cluster lifecycles, data replication, RBAC, and workload placement across middleware platforms.
  • Design and maintain observability systems using OpenTelemetry, Splunk, Prometheus, and Grafana, including metrics, logs, alerts, SLIs, and SLOs.
  • Respond to production incidents, lead post-mortems, and drive rapid system recovery.
  • Reduce operational toil through automation, code, disaster recovery strategies, and continuous resiliency testing.
  • Lead technical discussions, collaborate across functions, and mentor junior SRE colleagues.

Requirements

  • Programming or scripting experience with Python, Go, Java, or Bash.
  • Proficiency with Linux, distributed systems, containerization, Kubernetes/GKE, and GCP/GCE.
  • Understanding of CI/CD and infrastructure-as-code tools such as Terraform, Ansible, or Kubernetes Config Connector.
  • Knowledge of TCP/IP, UDP, HTTP, DNS, load balancing, and messaging protocols.
  • Ability to troubleshoot complex system behavior and communicate technical requirements across teams.
  • Experience applying generative AI and agents such as Gemini to platform operations.

Nice to have

  • Hands-on experience with OpenTelemetry, Splunk, Prometheus, and Grafana.
  • Experience with Agile development practices and software development lifecycles.
  • GCP Professional Cloud Architect, CKA, or CKAD certification.
  • Experience in financial markets or other regulated, ultra-low-latency, high-concurrency environments.

Culture & Benefits

  • Code-first engineering culture focused on systematic automation.
  • Bonus, equity, and employee stock purchase programmes.
  • Private medical and dental coverage, mental health benefits, pension, income protection, and life assurance.
  • Family leave, education assistance, and ongoing development and certification training.
  • Cycle-to-work scheme, EV car benefit scheme, and gym membership.
  • Hybrid working arrangement.

Π‘ΡƒΠ΄ΡŒΡ‚Π΅ остороТны: Ссли Ρ€Π°Π±ΠΎΡ‚ΠΎΠ΄Π°Ρ‚Π΅Π»ΡŒ просит Π²ΠΎΠΉΡ‚ΠΈ Π² ΠΈΡ… систСму, ΠΈΡΠΏΠΎΠ»ΡŒΠ·ΡƒΡ iCloud/Google, ΠΏΡ€ΠΈΡΠ»Π°Ρ‚ΡŒ ΠΊΠΎΠ΄/ΠΏΠ°Ρ€ΠΎΠ»ΡŒ, Π·Π°ΠΏΡƒΡΡ‚ΠΈΡ‚ΡŒ ΠΊΠΎΠ΄/ПО, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡ‚Π΅ этого - это мошСнники. ΠžΠ±ΡΠ·Π°Ρ‚Π΅Π»ΡŒΠ½ΠΎ ΠΆΠΌΠΈΡ‚Π΅ "ΠŸΠΎΠΆΠ°Π»ΠΎΠ²Π°Ρ‚ΡŒΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡˆΠΈΡ‚Π΅ Π² ΠΏΠΎΠ΄Π΄Π΅Ρ€ΠΆΠΊΡƒ. ΠŸΠΎΠ΄Ρ€ΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β†’