4 часа назад
Senior Infrastructure Engineer (AI)
168 000 - 213 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Infrastructure Engineer (AI) (Kubernetes, AWS, Terraform): Building secure, compliant infrastructure for containerized AI systems deployed across on-premises and cloud environments with an accent on Kubernetes operations, progressive delivery, observability, and infrastructure automation. Focus on architecting enterprise-grade deployments, hardening systems for regulated financial institutions, responding to high-severity incidents, and scaling low-latency AI inference.
Location: San Francisco, CA; on-site
Salary: $168,000–$213,000 per year plus equity
Company
provides an AI-native operations platform for regulated financial institutions, covering compliance, risk, fraud, and back-office operations using customer data and policies.
What you will do
- Own and evolve Kubernetes infrastructure, including cluster management, service mesh configuration, and container security policies.
- Design progressive delivery pipelines with canary deployments, automated rollbacks, and deployment health validation.
- Build and maintain Datadog observability infrastructure, including dashboards, monitors, SLOs, and distributed tracing.
- Lead incident response for high-severity outages and model capacity requirements for low-latency AI inference.
- Architect secure on-premises and cloud infrastructure using Infrastructure-as-Code, including VPCs, IAM policies, Kubernetes manifests, and private-cloud deployments.
- Lead enterprise customer rollouts, support SOC 2 infrastructure controls, and mentor mid-level engineers.
Requirements
- 8+ years of experience in infrastructure engineering or DevOps at high-growth or hyperscale companies.
- Production experience managing Docker and Kubernetes clusters, including Helm and service mesh technologies.
- Enterprise-scale experience with AWS, GCP, or Azure; AWS is preferred.
- Experience with Datadog or comparable observability platforms, including metrics, logs, APM, and distributed tracing.
- Strong Infrastructure-as-Code and safe deployment experience with Terraform, Helm, Kustomize, GitOps, progressive delivery, and automated rollbacks.
- Strong Python programming skills and experience leading outages, capacity events, and incident reviews.
Nice to have
- Familiarity with TypeScript.
- Experience preparing for or remediating SOC 2 or other compliance audits.
- Experience with private-cloud or on-premises deployments for regulated customers.
- Startup experience scaling infrastructure from early stages to enterprise level.
- Fintech or regulated-industry experience, including AI/ML infrastructure and model deployment at scale.
Culture & Benefits
- Work at the intersection of technical leadership, infrastructure as a product, and customer impact.
- Comprehensive healthcare, 401(k) matching, and commuter benefits.
- 15 days of paid time off plus holidays and unlimited sick days.
- Flexible leave options.
- DoorDash meals and Uber home coverage when working late.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 часов назад
Infrastructure Engineer (AI)
160 000 - 245 000$
13 часов назад
Staff Software Engineer (AWS/Kubernetes)
175 000 - 245 000$
17 часов назад
DevOps Engineer (AI)
85 000 - 180 000$
13 часов назад
Senior Software Engineer (AWS/Kubernetes)
160 000 - 215 000$
9 часов назад
Infrastructure Engineer
140 000 - 260 000$
12 часов назад
Senior Infrastructure Engineer (AI)
120 000 - 200 000$