Назад

Senior DevOps / SRE Support Engineer (Kubernetes)

5 000 - 5 500$
Формат работы
remote (только Latam)
Тип работы
fulltime
Грейд
senior
Английский
b2
vacancy_detail.hirify_telegram_tooltipВакансия из Telegram канала -

Мэтч & Сопровод

Покажет вашу совместимость и напишет письмо

Описание вакансии

TL;DR
Senior DevOps / SRE Support Engineer (Kubernetes): Supporting a Virtual NOC with production incident response, automation, observability, and AI-assisted triage for international client environments with an accent on Kubernetes, cloud infrastructure, and SLA-driven operations. Focus on resolving Sev1–Sev4 incidents, developing runbooks and automation, tuning Datadog/Prometheus/Grafana, and contributing to RAG-based L1 support tooling.

Senior DevOps / SRE Support Engineer

Прямой работодатель ( )

Сеньор Информационные технологии • DevOps • Kubernetes • Java • SQL • AWS • Azure • Google Cloud • ML/AI 22 сентября Удаленная работа
Опыт работы более 5 лет от 5 000 до 5 500 $ Работодатель Короткая ссылка: geekjob.ru/hijO Описание вакансии


Senior DevOps / SRE Support Engineer Virtual NOC (LATAM)

📍 LATAM only (AMER hub, GMT-3…-6) · 🕘 09:00–18:00 local · 💼 Full-time · 🗣 Advanced English · 💸 B2B, Up to 5500 $/monthly Our client is building a Virtual NOC — a managed DevOps/SRE practice providing 24/7 SLA-driven support for international clients on a follow-the-sun model. We are hiring a Senior DevOps / SRE Support Engineer to be among the founding members of the client's AMER hub.

What you'll do

  • Production incident response and resolution (Sev1–Sev4)
  • On-call duty rotation (1 week per 5-week cycle, separate compensation)
  • Runbook authoring (managed as code with CI validation)
  • Automation development and infrastructure changes for client environments
  • Observability tuning (Datadog, Prometheus, Grafana)
  • Post-mortem authoring and customer-facing technical communication
  • Knowledge base contribution
  • 30–40% of time allocated to engineering projects (automation, AI tooling, observability)
  • Contribute to AI-augmented L1 layer (RAG-based triage and automation)

Must-have

  • 5+ years hands-on in: Kubernetes, Terraform, AWS/GCP/Azure, Datadog, Prometheus, Grafana, PagerDuty, Linux, Python, Bash, ArgoCD, Helm, CI/CD (GitLab CI / GitHub Actions)
  • Production on-call experience — required
  • Java developer or SRE background
  • Ability to work with AI agents (critical thinking)
  • Reading foreign code and making small fixes
  • SQL
  • Incident management (Sev1–Sev4)
  • Advanced English — mandatory for direct client communication

Soft skills

  • Attention to detail · Responsibility · Strong communication

Location

  • LATAM only (AMER hub, GMT-3…-6)

Special conditions

  • Standard 09:00–18:00 local working hours covering AMER business hours, fully remote
  • On-call duty: 1 week per rotation with separate compensation (stand-by + per-incident)
  • No night shifts; mandatory recovery time after on-call

Why join

  • Founding member of the AMER hub
  • Real autonomy — not a ticket-processing position
  • 30–40% of time on engineering projects
  • Work directly with the Tech Lead and CEMEA team

Специализация

Отрасль и сфера применения

Уровень должности


Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →

Текст вакансии взят без изменений

Источник -