Назад
Company hidden
3 дня назад

Site Reliability Engineer

100 000 - 130 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
junior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (Cloud Infrastructure/AI): Building and supporting reliable, scalable infrastructure for AI-powered content classification products with an accent on multi-cloud systems, Kubernetes, infrastructure as code, and CI/CD. Focus on improving observability and production health, debugging application and infrastructure issues, and participating in a 24/7 on-call rotation.

Location: Marina del Rey, California, United States; hybrid work

Salary: $100,000–$130,000 annually

Company

hirify.global develops AI-powered content classification, activation, and verification solutions for brands and advertisers on platforms including YouTube, TikTok, Meta, and Snap.

What you will do

  • Build and support systems that help engineers deploy and manage product features safely and efficiently.
  • Deploy and maintain multi-cloud microservices using GitHub Actions, Argo CD, and Kubernetes.
  • Write and maintain infrastructure as code with Terraform and Terragrunt through pull requests and code review.
  • Build and improve CI/CD pipelines, release workflows, monitoring, and production health processes.
  • Debug application- and infrastructure-level issues and maintain documentation and runbooks.
  • Participate in a 24/7 on-call rotation alongside senior SRE teammates.

Requirements

  • 1–3 years of experience supporting cloud infrastructure in production with AWS and/or Google Cloud.
  • Hands-on experience with containers and Kubernetes.
  • Competency in Python and shell scripting.
  • Familiarity with CI/CD pipelines and GitOps, including GitHub Actions, GitLab, or Argo CD.
  • Exposure to monitoring and observability tools such as Prometheus, Grafana, Chronosphere, Datadog, or OpenTelemetry.
  • Strong problem-solving, communication, organization, documentation, and cross-functional collaboration skills.

Nice to have

  • Experience with GCP, AWS, Terraform, Terragrunt, Docker, FastAPI, Flask, Node, React, Airflow, Ray, PostgreSQL, DynamoDB, OpenSearch, or Snowflake.

Culture & Benefits

  • Flexible paid time off, 13 paid holidays, and Summer Fridays.
  • Medical, dental, and vision insurance with FSA options.
  • Company-paid life insurance and paid parental leave.
  • 401(k) with company match.
  • Professional development opportunities.
  • In-office lunches, free food, and optional in-person and virtual events.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →