Назад
Company hidden
20 часов Π½Π°Π·Π°Π΄

Site Reliability Engineer (SRE)

100Β 000 - 180Β 000$
Π€ΠΎΡ€ΠΌΠ°Ρ‚ Ρ€Π°Π±ΠΎΡ‚Ρ‹
remote (Ρ‚ΠΎΠ»ΡŒΠΊΠΎ USA)
Π’ΠΈΠΏ Ρ€Π°Π±ΠΎΡ‚Ρ‹
fulltime
Π“Ρ€Π΅ΠΉΠ΄
senior
Английский
b2
Π‘Ρ‚Ρ€Π°Π½Π°
US
Вакансия ΠΈΠ· списка Hirify.GlobalВакансия ΠΈΠ· Hirify Global, списка ΠΌΠ΅ΠΆΠ΄ΡƒΠ½Π°Ρ€ΠΎΠ΄Π½Ρ‹Ρ… tech-ΠΊΠΎΠΌΠΏΠ°Π½ΠΈΠΉ
Для мэтча ΠΈ ΠΎΡ‚ΠΊΠ»ΠΈΠΊΠ° Π½ΡƒΠΆΠ΅Π½ Plus

ΠœΡΡ‚Ρ‡ & Π‘ΠΎΠΏΡ€ΠΎΠ²ΠΎΠ΄

Для мэтча с этой вакансиСй Π½ΡƒΠΆΠ΅Π½ Plus

ОписаниС вакансии

ВСкст:
/
TL;DR
Site Reliability Engineer (SRE) (Kubernetes/Cloud): Operating and improving large-scale distributed systems in production with an accent on observability, incident response, and infrastructure automation. Focus on designing reliable Kubernetes platforms, building CI/CD and monitoring systems, conducting capacity and performance engineering, and strengthening resilience through chaos engineering and failure-mode analysis.

Location: 100% remote within the United States

Salary: $100,000–$180,000 annually

Company

hirify.global is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

What you will do

  • Define and refine SLOs, SLIs, and error budgets for critical production services.
  • Lead incident response, act as an incident commander when needed, and conduct post-incident reviews.
  • Design monitoring, logging, and tracing solutions using Prometheus, Grafana, OpenTelemetry, ELK/EFK, Datadog, or similar tools.
  • Automate operational workflows with Python, Go, Bash, or similar languages.
  • Architect and operate Kubernetes clusters, container workloads, autoscaling, capacity planning, network policies, and service-mesh integrations.
  • Build CI/CD pipelines and improve reliability through canary releases, chaos engineering, failover design, security improvements, and developer enablement.

Requirements

  • 10+ years of SRE, DevOps, or production engineering experience with large-scale distributed systems.
  • Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
  • Strong programming skills in Python, Go, or Java.
  • Deep hands-on Linux experience, including networking, performance tuning, and systems troubleshooting.
  • Production experience with Kubernetes, containers, observability tooling, and CI/CD pipelines.
  • Experience with distributed-system design, incident response, documentation, and technical communication.

Nice to have

  • Experience with production SLOs, error budgets, chaos engineering, capacity planning, performance engineering, or large-scale load testing.
  • Experience with AWS, Azure, or GCP.
  • Familiarity with Istio, Linkerd, Consul, Chaos Monkey, Gremlin, or Litmus.

Culture & Benefits

  • Full-time direct W2 employment.
  • Career growth opportunities within an established organization.
  • Blameless culture focused on operational excellence.
  • Mentorship and collaboration with application development and security teams.
  • New H-1B visa petitions are not sponsored; U.S. citizens, Green Card holders, EAD holders, and H-1B transfer candidates are encouraged to apply.

Π‘ΡƒΠ΄ΡŒΡ‚Π΅ остороТны: Ссли Ρ€Π°Π±ΠΎΡ‚ΠΎΠ΄Π°Ρ‚Π΅Π»ΡŒ просит Π²ΠΎΠΉΡ‚ΠΈ Π² ΠΈΡ… систСму, ΠΈΡΠΏΠΎΠ»ΡŒΠ·ΡƒΡ iCloud/Google, ΠΏΡ€ΠΈΡΠ»Π°Ρ‚ΡŒ ΠΊΠΎΠ΄/ΠΏΠ°Ρ€ΠΎΠ»ΡŒ, Π·Π°ΠΏΡƒΡΡ‚ΠΈΡ‚ΡŒ ΠΊΠΎΠ΄/ПО, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡ‚Π΅ этого - это мошСнники. ΠžΠ±ΡΠ·Π°Ρ‚Π΅Π»ΡŒΠ½ΠΎ ΠΆΠΌΠΈΡ‚Π΅ "ΠŸΠΎΠΆΠ°Π»ΠΎΠ²Π°Ρ‚ΡŒΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡˆΠΈΡ‚Π΅ Π² ΠΏΠΎΠ΄Π΄Π΅Ρ€ΠΆΠΊΡƒ. ΠŸΠΎΠ΄Ρ€ΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β†’