Назад
Company hidden
обновлено 5 дней назад

Senior DevOps Engineer (AI)

160 000 - 180 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior DevOps Engineer (AI) (Kubernetes/HPC): Building reliable, secure, and cost-efficient cloud infrastructure, self-service platforms, and AI-enabled operational tooling for a weather intelligence platform with an accent on Kubernetes, HPC clusters, and GPU-based model deployment. Focus on integrating MLOps, supporting scientific computing at scale, improving developer independence, and maintaining production availability through on-call operations.

Location: Washington, District of Columbia, United States. Access to technology controlled under U.S. export-control laws restricts the position to U.S. citizens, permanent residents, and protected individuals unless required licenses are obtained.

Salary: $160,000–$180,000 annually, depending on local market, skills, and experience.

Company

hirify.global develops a weather resilience platform combining space technology, generative AI, and proprietary weather modeling for forecasting and operational decision-making.

What you will do

  • Develop AI-powered tools and agentic workflows that improve development and operations efficiency.
  • Build and evolve adaptive cloud infrastructure focused on security, cost efficiency, SLOs, reliability, and developer experience.
  • Create self-service platforms for developers and weather scientists.
  • Operate Kubernetes-based product services and support scientific workloads on HPC clusters using SLURM.
  • Introduce MLOps practices for GPU-based model deployment on Kubernetes.
  • Collaborate with engineers, weather scientists, and spacecraft mission operations teams while participating in production on-call rotations.

Requirements

  • At least 6 years of experience as a Platform, DevOps, or SRE Engineer in containerized cloud environments.
  • Hands-on experience with AWS, GCP, or Azure and infrastructure as code such as Terraform or Crossplane.
  • Daily use of AI coding agents such as Claude Code or Copilot is required.
  • Experience with Kubernetes CI/CD, deployment methodologies, and monitoring systems such as Datadog, Prometheus, Grafana, or ELK Stack.
  • Proficiency in Python, Node, and Go, plus strong ownership, communication, and problem-solving skills.
  • Must meet U.S. export-control work authorization requirements.

Nice to have

  • Experience with HPC and scientific computing, including SLURM, AWS ParallelCluster, or Azure CycleCloud.
  • Familiarity with parallel filesystems such as Lustre or NFS.
  • Experience building agentic DevOps workflows.

Culture & Benefits

  • Work in a small global team with significant ownership and opportunities for growth.
  • Flexible working hours and unlimited paid time off.
  • Comprehensive health benefits.
  • Culture centered on impact, transparency, ownership, empathy, collaboration, and continuous growth.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →