Назад
Company hidden
1 день назад

Senior Staff Engineer (SRE)

120 000 - 260 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Staff Engineer (SRE): Building and operating incident-management automation, shared services, APIs, dashboards, and data pipelines for high-availability distributed platforms with an accent on observability, reliability engineering, and production incident response. Focus on designing resilient cloud-native systems, leading high-severity incident recovery, and turning root-cause analysis into systemic engineering improvements.

Location: Bethesda, Maryland, United States

Annual salary: $120,000–$260,000

Company

hirify.global is a large United States auto insurer and a member of the Berkshire Hathaway family of companies, serving millions of customers nationwide.

What you will do

  • Design, develop, and operate automation, self-service tools, dashboards, data pipelines, APIs, and shared services for incident management and on-call operations.
  • Improve detection, troubleshooting, recovery, runbooks, paging, and post-incident processes across engineering teams.
  • Lead technical response during high-severity incidents, coordinating troubleshooting, impact analysis, and safe service restoration.
  • Lead root-cause analysis, post-incident reviews, corrective actions, and systemic reliability improvements.
  • Set engineering standards for architecture, deployment, testing, observability, security, and production readiness.
  • Partner with SRE, platform, infrastructure, security, product, and business stakeholders while mentoring senior engineers and teams.

Requirements

  • 10+ years of professional software engineering experience and 8+ years in architecture, system reliability, scalability, or technical leadership.
  • Hands-on experience with multiple languages, including Go, Java, Python, and C#, and production systems running on Kubernetes and serverless technologies.
  • Experience with Azure, AWS, or other cloud platforms; SQL and NoSQL technologies; data pipelines; analytics; and operational dashboards.
  • Experience with OpenTelemetry, Grafana, Datadog, Splunk, Azure Monitor, PagerDuty, and incident-management or observability platforms.
  • Strong incident forensics, root-cause analysis, distributed-systems design, reliability engineering, and technical communication skills.
  • Must participate in a 24x7 on-call rotation and support high-severity production incidents.

Nice to have

  • Experience with platform engineering, backend engineering, distributed systems, operational tooling, or complex hybrid environments.
  • Bachelor’s degree in Computer Science, Information Systems, or equivalent education or work experience.
  • Experience with AI-assisted development tools such as Claude Code, Cursor, and GitHub Copilot.

Culture & Benefits

  • Personalized development programs, mentorship, and certification assistance.
  • Inclusive and collaborative culture focused on shared success and continuous improvement.
  • Competitive pay, benefits, and flexibility supporting employee well-being.
  • Engineers follow a “You Build It, You Run It” model with ownership of the systems they build and operate.

Hiring process

  • Selection considers role scope, experience, education, training, work location, and business factors.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →