Назад
Company hidden
7 часов назад

AI Engineer - Cloud Infrastructure (AI)

175 000 - 275 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Engineer - Cloud Infrastructure (AI): Designing and operating Terraform-defined infrastructure and Kubernetes environments that power highly available, multi-tenant enterprise AI products with an accent on AWS, cloud networking, autoscaling, and security. Focus on building resilient infrastructure for AI workloads, improving observability and CI/CD, and solving complex reliability, capacity, and incident-response challenges.

Location: New York, United States; full-time, in-person, 5 days a week

Salary: $175,000–$275,000 USD per year, plus equity and benefits

Company

hirify.global builds AI-powered site reliability software that troubleshoots, remediates, and helps prevent complex production incidents for enterprise customers.

What you will do

  • Design scalable, reliable infrastructure for AI workloads, inference, data pipelines, and agentic workflows.
  • Build CI/CD and software development lifecycle tooling for the engineering organization.
  • Operate and improve highly available, resilient, and cost-efficient AWS and Kubernetes platforms.
  • Evolve Terraform and Helm deployments across multiple environments, including secrets, policy-as-code, and workload identity.
  • Build end-to-end observability and connect infrastructure signals to the AI SRE agent.
  • Improve security and compliance through least privilege, just-in-time access, default-deny egress, auditability, and policy-as-code.

Requirements

  • 7+ years of experience in technically rigorous engineering environments.
  • Hands-on experience operating AWS and Kubernetes-native infrastructure and applications at scale with more than 99.9% availability.
  • Practical experience with AWS, EKS, Terraform, and Helm.
  • Experience designing idempotent systems, including outbox patterns, deduplication keys, and safe replay.
  • Experience with incident response, chaos testing, capacity planning, and debugging across infrastructure, compute, networking, runtime, storage, and authentication layers.
  • Ability to work in person in New York five days per week.

Nice to have

  • Experience with Envoy, Istio, Cilium, or eBPF.
  • Experience operating GPU workloads, inference servers, or token-streaming gateways.
  • Production experience with Python, Rust, and TypeScript.
  • Experience with data governance, PII discovery and redaction, lineage, or tokenization.
  • Experience designing cross-region active/active architectures or working with GCP, Azure, or Oracle Cloud.

Culture & Benefits

  • Health insurance and additional employee benefits.
  • Startup equity and competitive compensation.
  • Flexible time off.
  • Technical equipment and in-office snacks.
  • Collaborative, high-impact environment focused on AI-powered software maintenance.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →