Назад
Company hidden
6 дней назад

Cloud Operations Manager (AWS)

Формат работы
remote
Тип работы
fulltime
Грейд
lead
Английский
b2
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Cloud Operations Manager (AWS): Leading the IT operations team and elevating cloud, GPU compute, and edge infrastructure maturity with an accent on operational rigor, CI/CD automation, and AWS architecture. Focus on implementing scalable infrastructure patterns, mentoring technical staff, and establishing robust incident response and change management processes.

Location: Remote

Company

hirify.global is a technology company focused on cloud, GPU compute, and edge device infrastructure.

What you will do

  • Manage and mentor the IT operations team, providing technical guidance and hands-on coaching.
  • Own and scale the AWS environment, implementing architecture patterns, cost optimization, and security hardening.
  • Design and maintain automated CI/CD pipelines for containerized, hardware-dependent, and ML-related services.
  • Build operational maturity through documentation, runbooks, and standardized change management processes.
  • Partner with engineering teams to provide reliable infrastructure that enables fast and safe deployments.
  • Lead incident response efforts and conduct blameless post-mortems to reduce recurring system issues.

Requirements

  • 5+ years of hands-on experience with AWS (EC2, RDS, S3, ECS, Fargate, IAM, IaC, networking).
  • Strong background in CI/CD pipeline design and deployment automation using tools like GitHub Actions or Jenkins.
  • Proven experience managing and developing technical team members through coaching and mentorship.
  • Track record of implementing operational processes such as on-call rotations and incident response.
  • Solid understanding of cloud security best practices, IAM policies, and SOC 2 considerations.
  • Ability to communicate technical complexity clearly to both engineering and non-technical stakeholders.

Nice to have

  • Experience optimizing costs for ML/GPU workloads.
  • Proficiency with Terraform or CloudFormation.
  • Knowledge of Kubernetes and container orchestration.
  • Familiarity with monitoring tools like DataDog, CloudWatch, or PagerDuty.
  • Scripting skills in Python, Bash, or Ruby.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →