Назад
Company hidden
17 часов назад

Infrastructure Specialist (Kubernetes)

125 000 - 130 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Infrastructure Specialist (Kubernetes): Operating and improving the cloud infrastructure, Kubernetes clusters, observability platform, and managed Airflow service with an accent on reliability, customer troubleshooting, and multi-cloud operations. Focus on building monitoring and alerting systems, automating operational tasks, responding to incidents, and permanently resolving complex distributed-system issues.

Location: Remote within the United States; Austin, Boston, or New York City

Estimated total compensation: $125,000–$130,000 per year, plus equity and comprehensive benefits.

Company

hirify.global develops Astro, a unified DataOps platform powered by Apache Airflow for reliable data products, analytics, AI, and data-driven applications.

What you will do

  • Operate, monitor, and maintain cloud infrastructure and Kubernetes clusters supporting the managed Airflow service.
  • Provide technical solutions, troubleshoot customer environments, and lead incident triage with customers.
  • Participate in the weekend on-call rotation, meet SLAs, and permanently resolve or monitor production issues.
  • Build monitoring, alerting, and operational automation to improve platform reliability.
  • Provide customer feedback to product development teams and contribute to product architecture.
  • Improve customer documentation and provide guidance through the path to production.

Requirements

  • 5 years of experience with large, complex cloud infrastructures operating at scale.
  • 3 years of experience with Kubernetes.
  • Production distributed-systems experience with AWS, GCP, or Azure.
  • Strong Linux, troubleshooting, distributed-systems monitoring, and customer-support experience.
  • DevOps or CI/CD experience and Python scripting skills.
  • Strong communication skills and willingness to participate in weekend on-call coverage.

Nice to have

  • Site Reliability Engineer experience.
  • Experience with Kubernetes Custom Resources, Azure, Airflow or big-data orchestration, and infrastructure as code.

Culture & Benefits

  • Work remotely as part of a fully distributed team.
  • Work with customers across different industries, cloud providers, and technical environments.
  • Use current technologies and multi-cloud implementations.
  • Receive equity and a comprehensive benefits package.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →