16 часов назад
Principal Site Reliability Engineer (Temp to Hire) (SRE)
165 000 - 185 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Site Reliability Engineer (SRE): Building reliable, available, and performant production platforms with an accent on incident management, observability, infrastructure automation, and compliance readiness. Focus on designing SLOs and on-call operations, automating Terraform and CI/CD workflows, testing disaster recovery, and reducing operational toil across distributed teams.
Location: Fully remote within the United States. Training is virtual and equipment is provided.
Salary: $165,000–$185,000 annually, plus bonus and benefits.
Company
develops and manufactures diabetes technology, including the Tandem Mobi system and t:slim X2 insulin pump with Control-IQ+ technology.
What you will do
- Lead production support, incident command, escalation, stakeholder communication, and blameless postmortems.
- Define SLIs and SLOs, improve observability, tune alerts, reduce MTTD and MTTR, and build operational runbooks.
- Own on-call strategy, rotation design, escalation paths, and coverage across a distributed team.
- Automate infrastructure with Terraform and improve CI/CD reliability through rollback, risk-check, and progressive-delivery guardrails.
- Lead business continuity and disaster recovery readiness, including backups, failover, recovery testing, and RTO/RPO validation.
- Partner with Security, Quality, Compliance, software engineering, QA, and architecture while mentoring internal and consulting-partner engineers.
Requirements
- Experience leading production support, high-severity incident management, and postmortem practices.
- Strong knowledge of SRE principles, including SLIs/SLOs, blameless postmortems, toil reduction, and reliability engineering.
- Expertise with Terraform or comparable infrastructure as code at scale, including modules, state management, and policy guardrails.
- Hands-on experience with CI/CD pipelines, a major cloud platform, Docker, Kubernetes, observability tooling, and scripting or programming such as Python, Go, or Bash.
- Experience designing and testing disaster recovery, cloud security, compliance controls, and cost optimization.
- 10+ years in SRE, DevOps, or infrastructure engineering and 2+ years mentoring or technically leading remote, offshore, or contracted engineers.
Nice to have
- Experience in FDA- and ISO-regulated industries and with agile methodologies.
- B.S. in Computer Science or equivalent education and production experience.
- Professional- or architect-level AWS, Azure, or GCP certification.
Culture & Benefits
- Temporary assignment through a staffing agency, with potential conversion to regular full-time employment based on business need and performance.
- Agency benefits may be available during the temporary assignment; Tandem-sponsored benefits are available after conversion.
- Remote work with virtual training and company-provided equipment.
- Collaborative, inclusive workplace focused on innovation, growth, well-being, and camaraderie.
Hiring process
- The position remains open until a final candidate is selected or sufficient applications are received.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 часов назад
Site Reliability Engineer (Cybersecurity)
96 500 - 183 500$
2 дня назад
Site Reliability Engineer (AWS GovCloud)
91 100 - 125 200$
Okta
5 дней назад
Staff Site Reliability Engineer (FedRAMP)
194 000 - 267 000$
3 дня назад
Senior Site Reliability Engineer (Fintech)
160 000 - 200 000$
2 дня назад
Software Development Engineer, SRE (US Federal)
137 000 - 205 400$
5 дней назад
Sr. Site Reliability Engineer (Healthcare)
125 000 - 145 000$