Назад
Company hidden
4 часа назад

Senior Infrastructure Engineer (AI)

168 000 - 213 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Infrastructure Engineer (AI) (Kubernetes, AWS, Terraform): Building secure, compliant infrastructure for containerized AI systems deployed across on-premises and cloud environments with an accent on Kubernetes operations, progressive delivery, observability, and infrastructure automation. Focus on architecting enterprise-grade deployments, hardening systems for regulated financial institutions, responding to high-severity incidents, and scaling low-latency AI inference.

Location: San Francisco, CA; on-site

Salary: $168,000–$213,000 per year plus equity

Company

hirify.global provides an AI-native operations platform for regulated financial institutions, covering compliance, risk, fraud, and back-office operations using customer data and policies.

What you will do

  • Own and evolve Kubernetes infrastructure, including cluster management, service mesh configuration, and container security policies.
  • Design progressive delivery pipelines with canary deployments, automated rollbacks, and deployment health validation.
  • Build and maintain Datadog observability infrastructure, including dashboards, monitors, SLOs, and distributed tracing.
  • Lead incident response for high-severity outages and model capacity requirements for low-latency AI inference.
  • Architect secure on-premises and cloud infrastructure using Infrastructure-as-Code, including VPCs, IAM policies, Kubernetes manifests, and private-cloud deployments.
  • Lead enterprise customer rollouts, support SOC 2 infrastructure controls, and mentor mid-level engineers.

Requirements

  • 8+ years of experience in infrastructure engineering or DevOps at high-growth or hyperscale companies.
  • Production experience managing Docker and Kubernetes clusters, including Helm and service mesh technologies.
  • Enterprise-scale experience with AWS, GCP, or Azure; AWS is preferred.
  • Experience with Datadog or comparable observability platforms, including metrics, logs, APM, and distributed tracing.
  • Strong Infrastructure-as-Code and safe deployment experience with Terraform, Helm, Kustomize, GitOps, progressive delivery, and automated rollbacks.
  • Strong Python programming skills and experience leading outages, capacity events, and incident reviews.

Nice to have

  • Familiarity with TypeScript.
  • Experience preparing for or remediating SOC 2 or other compliance audits.
  • Experience with private-cloud or on-premises deployments for regulated customers.
  • Startup experience scaling infrastructure from early stages to enterprise level.
  • Fintech or regulated-industry experience, including AI/ML infrastructure and model deployment at scale.

Culture & Benefits

  • Work at the intersection of technical leadership, infrastructure as a product, and customer impact.
  • Comprehensive healthcare, 401(k) matching, and commuter benefits.
  • 15 days of paid time off plus holidays and unlimited sick days.
  • Flexible leave options.
  • DoorDash meals and Uber home coverage when working late.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →