Назад
Company hidden
2 дня назад

AI Platform Operations Manager (Azure)

128 260 - 146 017$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Platform Operations Manager (Azure): Automating and operating infrastructure, CI/CD pipelines, container platforms, and observability for enterprise AI solutions with an accent on Azure infrastructure-as-code, AKS, Azure Container Apps, and MLOps operations. Focus on deploying agentic AI services and RAG pipelines, managing model and prompt releases, improving reliability, implementing DevSecOps controls, and optimizing AI infrastructure costs.

Location: Denver, Colorado, United States. Full-time role with less than 10% travel. Must be eligible to work in the United States and pass comprehensive background and drug screening.

Salary: $128,260.00–$146,017.59 per year.

Company

hirify.global provides wholesale, colocation, and cloud data centers for hyperscale and enterprise companies, focusing on availability, security, connectivity, and physical resilience.

What you will do

  • Build and maintain Azure infrastructure-as-code modules and reusable deployment patterns for AI platform services, including Azure AI Foundry, Azure OpenAI Service, Azure AI Search, Cosmos DB, ADLS Gen2, and Databricks.
  • Design and operate CI/CD pipelines for application code, infrastructure, containers, and AI agent deployments using Azure DevOps or GitHub Actions.
  • Deploy and operate agentic AI services, MCP servers, RAG pipelines, and integration workloads on Azure Container Apps and AKS.
  • Operate container platforms, including cluster upgrades, autoscaling, networking, workload isolation, GPU-backed workloads, and Azure Container Registry governance.
  • Implement observability, dashboards, alerting, runbooks, incident response, and automated remediation for platform and AI services.
  • Apply DevSecOps, governance, data residency, compliance, FinOps, and cost-optimization practices while partnering with AI, data, application, and security teams.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.
  • 5+ years of hands-on experience in DevOps, site reliability engineering, or platform engineering.
  • Strong Infrastructure as Code experience with Terraform, Bicep, or ARM, including module design and state management.
  • Experience building CI/CD pipelines, working with Docker and AKS, and operating Azure Container Apps or equivalent platforms.
  • Working knowledge of Azure compute, networking, storage, Entra ID, Key Vault, monitoring, observability, Git workflows, and artifact or registry management.
  • Strong scripting and automation skills in Python, PowerShell, or Bash, with proven production troubleshooting experience.

Nice to have

  • Microsoft Azure DevOps, Azure Administrator, or Certified Kubernetes Administrator certification.
  • Experience operating AI/ML workloads, model endpoints, RAG pipelines, vector databases, agentic services, or MLOps platforms.
  • Experience with MCP servers, Semantic Kernel, LangGraph, AutoGen, GPU provisioning, or inference cost optimization.
  • Knowledge of FinOps and experience integrating enterprise systems such as Microsoft 365, Workday, NetSuite, or Procore.

Culture & Benefits

  • Healthcare, dental, vision, life insurance, paid time off, and paid leave programs.
  • 401(k), flexible spending accounts, and a cell phone subsidy.
  • Peer-to-peer recognition and rewards programs.
  • Company events, game nights, happy hours, and barbecues.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →