Назад
Company hidden
16 часов назад

Principal Site Reliability Engineer (Temp to Hire) (SRE)

165 000 - 185 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal Site Reliability Engineer (SRE): Building reliable, available, and performant production platforms with an accent on incident management, observability, infrastructure automation, and compliance readiness. Focus on designing SLOs and on-call operations, automating Terraform and CI/CD workflows, testing disaster recovery, and reducing operational toil across distributed teams.

Location: Fully remote within the United States. Training is virtual and equipment is provided.

Salary: $165,000–$185,000 annually, plus bonus and benefits.

Company

hirify.global develops and manufactures diabetes technology, including the Tandem Mobi system and t:slim X2 insulin pump with Control-IQ+ technology.

What you will do

  • Lead production support, incident command, escalation, stakeholder communication, and blameless postmortems.
  • Define SLIs and SLOs, improve observability, tune alerts, reduce MTTD and MTTR, and build operational runbooks.
  • Own on-call strategy, rotation design, escalation paths, and coverage across a distributed team.
  • Automate infrastructure with Terraform and improve CI/CD reliability through rollback, risk-check, and progressive-delivery guardrails.
  • Lead business continuity and disaster recovery readiness, including backups, failover, recovery testing, and RTO/RPO validation.
  • Partner with Security, Quality, Compliance, software engineering, QA, and architecture while mentoring internal and consulting-partner engineers.

Requirements

  • Experience leading production support, high-severity incident management, and postmortem practices.
  • Strong knowledge of SRE principles, including SLIs/SLOs, blameless postmortems, toil reduction, and reliability engineering.
  • Expertise with Terraform or comparable infrastructure as code at scale, including modules, state management, and policy guardrails.
  • Hands-on experience with CI/CD pipelines, a major cloud platform, Docker, Kubernetes, observability tooling, and scripting or programming such as Python, Go, or Bash.
  • Experience designing and testing disaster recovery, cloud security, compliance controls, and cost optimization.
  • 10+ years in SRE, DevOps, or infrastructure engineering and 2+ years mentoring or technically leading remote, offshore, or contracted engineers.

Nice to have

  • Experience in FDA- and ISO-regulated industries and with agile methodologies.
  • B.S. in Computer Science or equivalent education and production experience.
  • Professional- or architect-level AWS, Azure, or GCP certification.

Culture & Benefits

  • Temporary assignment through a staffing agency, with potential conversion to regular full-time employment based on business need and performance.
  • Agency benefits may be available during the temporary assignment; Tandem-sponsored benefits are available after conversion.
  • Remote work with virtual training and company-provided equipment.
  • Collaborative, inclusive workplace focused on innovation, growth, well-being, and camaraderie.

Hiring process

  • The position remains open until a final candidate is selected or sufficient applications are received.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →