Назад
Company hidden
6 дней назад

Site Reliability Engineer III

130 000 - 153 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Site Reliability Engineer III (SRE/GCP): Build and maintain infrastructure that enables developers to ship reliably at scale, with an accent on infrastructure-as-code, deployment automation, and observability. Focus on incident response best practices, performance/availability/cost optimization, and integrating monitoring/alerting systems across production services.

Location: Bozeman, Montana, United States

Salary: $130,000–$153,000

Company

hirify.global builds technology products that help people confidently explore the outdoors.

What you will do

  • Deploy, monitor, and maintain highly available systems using Terraform, CockroachDB, and GCP services (including GKE, Cloud SQL, Bigtable, Composer/Airflow, Cloud Storage, BigQuery, Pub/Sub, and Cloud Run).
  • Maintain and extend a large, mature Terraform codebase and use infrastructure-as-code to keep a simple path to production.
  • Analyze systems and recommend improvements to performance, availability, and cost efficiency.
  • Automate manual processes to minimize toil and reduce operational burden.
  • Develop and maintain integrations with monitoring and alerting tools (Google Cloud Monitoring, Prometheus, OpenTelemetry, Checkly, Rootly).
  • Drive incident response best practices and participate in the SRE on-call rotation for core infrastructure.

Requirements

  • 5+ years of experience, with 3+ years supporting production systems.
  • B.S. or M.S. in computer science or a related field, or relevant experience.
  • Strong interest and experience with Kubernetes, networking, and infrastructure-as-code.
  • Experience with Terraform (or OpenTofu).
  • Exposure to at least one major cloud platform.
  • Practical experience with multiple datastore types (SQL, NoSQL, object storage) and ability to explain tradeoffs based on access patterns and scalability.

Nice to have

  • Familiarity with Google Cloud Platform.
  • Experience troubleshooting high-throughput, low-latency services.
  • Experience with IAM, auditing, and security management in a cloud environment.
  • Experience with GIS mapping systems and tiles.
  • Experience with Airflow (or equivalent ETL systems).

Culture & Benefits

  • Comprehensive health benefits, including a no-monthly-cost medical plan.
  • Parental leave plan of 5 or 13 weeks fully paid.
  • 401k matching at 100% for the first 3% saved and 50% from 3–5%.
  • Flexible time away package including PTO, STO, VTO, quiet weeks, and floating holidays.
  • Annual outdoor funds and company-wide outdoor adventures and perks.

Hiring process

  • Interviews and evaluation of technical fit for core infrastructure, reliability, and incident response.
  • Position is open until filled.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →