Назад
Company hidden
6 дней назад

Staff DevOps Engineer (AI)

Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff DevOps Engineer (AI): Building and operating AWS and Kubernetes infrastructure for an AI-powered mental healthcare platform with an accent on infrastructure automation, observability, security, and zero-downtime delivery. Focus on setting infrastructure strategy, integrating agentic AI into production workflows, and solving complex reliability, compliance, and developer-experience challenges.

Location: Fully remote; the role is offered through a B2B model.

Company

hirify.global operates an AI-powered mental healthcare platform that combines licensed providers with proprietary AI tools to deliver scalable psychiatric care in the United States.

What you will do

  • Set the technical vision and roadmap for company infrastructure, including measurable goals for reliability, cost, performance, deployment health, and developer velocity.
  • Run and evolve AWS and Kubernetes infrastructure, including EKS cluster management, autoscaling with Karpenter, policy enforcement with Kyverno, and zero-downtime operations.
  • Own Infrastructure as Code with Terraform and AWS CDK in TypeScript, as well as GitLab CI/CD templates, OIDC, and self-managed GitLab.
  • Build practical observability with Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM.
  • Maintain blue-green deployments, health-gated rollbacks, zero-downtime PostgreSQL migrations, and CI migration gating.
  • Partner with product engineering teams, lead security engineering in a HIPAA environment, and integrate autonomous-agent output into production workflows.

Requirements

  • 6+ years of DevOps or infrastructure engineering experience with strong Linux administration, troubleshooting, performance analysis, and systems fundamentals.
  • Hands-on experience with AWS services, production Kubernetes/EKS, cluster management, node scaling, policy enforcement, containers, networking, TLS, and DNS.
  • Strong Terraform, AWS CDK in TypeScript, GitLab CI/CD, reusable pipeline templates, OIDC, and self-managed GitLab experience.
  • Practical experience with monitoring, observability, secrets rotation, short-lived credentials, leak scanning, PHI-aware logging, and HashiCorp Vault as code.
  • Experience with blue-green deployments, automated rollback, PostgreSQL expand/contract migrations, Docker, ECR, immutable tags, and image lifecycle management.
  • Fluent Russian and English at B2 level; experience working effectively in remote, distributed teams.

Nice to have

  • Experience in HIPAA, SOC 2, or another regulated environment.
  • Ansible experience for VM fleet management.
  • Node.js and TypeScript application operations, including pm2 and npm.
  • GitOps tooling such as ArgoCD or Flux, deeper PostgreSQL administration, or AWS certifications.

Culture & Benefits

  • Fully remote long-term collaboration under a B2B model.
  • Health insurance after the probation period and sports and wellness compensation.
  • 19 paid vacation days, four additional wellness days, and paid sick leave for the first five working days.
  • Personalized English lessons via Preply and opportunities for training and career growth.
  • Supportive distributed culture with thoughtful gifts and offline corporate events.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →