7 часов назад
AI Engineer - Cloud Infrastructure (AI)
175 000 - 275 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
AI Engineer - Cloud Infrastructure (AI): Designing and operating Terraform-defined infrastructure and Kubernetes environments that power highly available, multi-tenant enterprise AI products with an accent on AWS, cloud networking, autoscaling, and security. Focus on building resilient infrastructure for AI workloads, improving observability and CI/CD, and solving complex reliability, capacity, and incident-response challenges.
Location: New York, United States; full-time, in-person, 5 days a week
Salary: $175,000–$275,000 USD per year, plus equity and benefits
Company
builds AI-powered site reliability software that troubleshoots, remediates, and helps prevent complex production incidents for enterprise customers.
What you will do
- Design scalable, reliable infrastructure for AI workloads, inference, data pipelines, and agentic workflows.
- Build CI/CD and software development lifecycle tooling for the engineering organization.
- Operate and improve highly available, resilient, and cost-efficient AWS and Kubernetes platforms.
- Evolve Terraform and Helm deployments across multiple environments, including secrets, policy-as-code, and workload identity.
- Build end-to-end observability and connect infrastructure signals to the AI SRE agent.
- Improve security and compliance through least privilege, just-in-time access, default-deny egress, auditability, and policy-as-code.
Requirements
- 7+ years of experience in technically rigorous engineering environments.
- Hands-on experience operating AWS and Kubernetes-native infrastructure and applications at scale with more than 99.9% availability.
- Practical experience with AWS, EKS, Terraform, and Helm.
- Experience designing idempotent systems, including outbox patterns, deduplication keys, and safe replay.
- Experience with incident response, chaos testing, capacity planning, and debugging across infrastructure, compute, networking, runtime, storage, and authentication layers.
- Ability to work in person in New York five days per week.
Nice to have
- Experience with Envoy, Istio, Cilium, or eBPF.
- Experience operating GPU workloads, inference servers, or token-streaming gateways.
- Production experience with Python, Rust, and TypeScript.
- Experience with data governance, PII discovery and redaction, lineage, or tokenization.
- Experience designing cross-region active/active architectures or working with GCP, Azure, or Oracle Cloud.
Culture & Benefits
- Health insurance and additional employee benefits.
- Startup equity and competitive compensation.
- Flexible time off.
- Technical equipment and in-office snacks.
- Collaborative, high-impact environment focused on AI-powered software maintenance.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 час назад
ML Platform Engineer (AI)
210 000 - 300 000$
9 часов назад
Principal Engineer (AI)
300 000 - 500 000$
1 день назад
Forward Deployed Engineer (Post-Sales) (AI)
230 000 - 300 000$
6 дней назад
Forward Deployed Engineer (AI)
180 000 - 260 000$
3 часа назад
Staff Software Engineer (AI)
180 000 - 240 000$
3 часа назад
Technical Staff
200 000 - 350 000$