3 дня назад
Site Reliability Engineer
100 000 - 130 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Cloud Infrastructure/AI): Building and supporting reliable, scalable infrastructure for AI-powered content classification products with an accent on multi-cloud systems, Kubernetes, infrastructure as code, and CI/CD. Focus on improving observability and production health, debugging application and infrastructure issues, and participating in a 24/7 on-call rotation.
Location: Marina del Rey, California, United States; hybrid work
Salary: $100,000–$130,000 annually
Company
develops AI-powered content classification, activation, and verification solutions for brands and advertisers on platforms including YouTube, TikTok, Meta, and Snap.
What you will do
- Build and support systems that help engineers deploy and manage product features safely and efficiently.
- Deploy and maintain multi-cloud microservices using GitHub Actions, Argo CD, and Kubernetes.
- Write and maintain infrastructure as code with Terraform and Terragrunt through pull requests and code review.
- Build and improve CI/CD pipelines, release workflows, monitoring, and production health processes.
- Debug application- and infrastructure-level issues and maintain documentation and runbooks.
- Participate in a 24/7 on-call rotation alongside senior SRE teammates.
Requirements
- 1–3 years of experience supporting cloud infrastructure in production with AWS and/or Google Cloud.
- Hands-on experience with containers and Kubernetes.
- Competency in Python and shell scripting.
- Familiarity with CI/CD pipelines and GitOps, including GitHub Actions, GitLab, or Argo CD.
- Exposure to monitoring and observability tools such as Prometheus, Grafana, Chronosphere, Datadog, or OpenTelemetry.
- Strong problem-solving, communication, organization, documentation, and cross-functional collaboration skills.
Nice to have
- Experience with GCP, AWS, Terraform, Terragrunt, Docker, FastAPI, Flask, Node, React, Airflow, Ray, PostgreSQL, DynamoDB, OpenSearch, or Snowflake.
Culture & Benefits
- Flexible paid time off, 13 paid holidays, and Summer Fridays.
- Medical, dental, and vision insurance with FSA options.
- Company-paid life insurance and paid parental leave.
- 401(k) with company match.
- Professional development opportunities.
- In-office lunches, free food, and optional in-person and virtual events.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 дней назад
Site Reliability Engineer
123 000 - 150 000$
9 дней назад
Senior Site Reliability Engineer (Fintech)
160 000 - 200 000$
Nebius
4 дня назад
Senior Site Reliability Engineer (AI)
147 200 - 224 000$
5 дней назад
Sr. Staff Site Reliability Engineer-Federal, Security Clearance
164 000 - 205 000$
10 дней назад
Senior Site Reliability Engineer (AWS)
130 000 - 155 000$
10 дней назад
Senior Site Reliability Engineer (AWS/AI)
140 000 - 160 000$