8 дней назад
Senior/Staff DevOps Engineer (AI)
100 100$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior/Staff DevOps Engineer (AI/cloud infrastructure): Building and operating AWS and Kubernetes infrastructure, infrastructure-as-code, CI/CD, observability, and security systems for an AI-powered mental healthcare platform with an accent on reliability, zero-downtime operations, HIPAA-aware security, and developer experience. Focus on designing scalable EKS environments, implementing automated deployment rollbacks and PostgreSQL migrations, and integrating agentic AI workflows into production.
Location: Fully remote long-term collaboration; no specific candidate location is stated.
Estimated annual earnings: $100,100. This is an estimate and is not a guaranteed minimum.
Company
is an AI-powered mental healthcare platform combining licensed providers with proprietary AI tools to deliver scalable psychiatric care in the United States.
What you will do
- Set the technical vision and roadmap for company infrastructure, defining measurable goals for reliability, cost, performance, deployment health, and developer velocity.
- Run and evolve AWS and Kubernetes infrastructure, including EKS cluster management, autoscaling with Karpenter, policy enforcement with Kyverno, and zero-downtime operations.
- Own infrastructure as code with Terraform and AWS CDK in TypeScript, together with GitLab CI/CD templates, OIDC, and self-managed GitLab.
- Build practical observability using Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM.
- Implement blue-green deployments, health-gated automated rollback, zero-downtime PostgreSQL migrations, and CI migration gating.
- Improve security and developer experience in a HIPAA environment while integrating autonomous-agent output into production workflows.
Requirements
- 6+ years of DevOps or infrastructure engineering experience with strong Linux administration, troubleshooting, performance analysis, and systems fundamentals.
- Hands-on AWS and production Kubernetes/EKS experience, including cluster management, node scaling, policy enforcement, Karpenter, or Kyverno.
- Strong experience with Terraform, AWS CDK in TypeScript, GitLab CI/CD, reusable templates, OIDC, and self-managed GitLab.
- Practical experience with observability, security engineering, secrets rotation, short-lived credentials, PHI-aware logging, and HashiCorp Vault.
- Experience with blue-green deployments, automated rollback, PostgreSQL zero-downtime migrations, Docker, ECR, load balancing, TLS, and DNS.
- Fluent Russian and English at B1 level required. Experience working effectively in remote, distributed teams and using agentic AI workflows is required.
Nice to have
- Experience in HIPAA, SOC 2, or another regulated environment.
- Ansible, Node.js operations, GitOps with ArgoCD or Flux, deeper PostgreSQL administration, or AWS certifications.
Culture & Benefits
- Remote-first collaboration under a B2B model with flexible long-term work arrangements.
- Health insurance after the probation period, plus sports and wellness compensation.
- 19 paid vacation days, 4 additional wellness days annually, and paid sick leave for the first 5 working days.
- Personalized English lessons via Preply and opportunities for career growth.
- Supportive distributed culture with gifts for key life events and offline corporate events.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →