2 часа назад
Senior Platform Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Platform Engineer (AI): Owning and scaling AWS, Kubernetes, infrastructure-as-code, CI/CD, and production systems for AI agents across customer environments with an accent on deployment orchestration, inference reliability, observability, and security. Focus on designing multi-tenant rollouts, optimizing GPU-backed model serving, and maintaining performance and reliability as customer environments and inference volume grow.
Location: Hybrid New York City model preferred; remote work considered for the right person
Company
provides a domain-specific AI platform for the Architecture, Engineering, and Construction industry, combining embedding models, document parsing, and autonomous agents to extract structured knowledge from enterprise project files.
What you will do
- Orchestrate agent and service rollouts across multiple customer cloud environments, including deployment strategies, per-customer configuration, automated health checks, and monitoring.
- Build and operate fast, reliable, and cost-effective inference infrastructure, including GPU workloads, serving systems, and model deployment.
- Own core infrastructure across multi-account AWS, Kubernetes, CI/CD, observability, disaster recovery, and cost management.
- Strengthen security through access controls, secrets management, network security, image scanning, dependency auditing, and SOC 2 compliance work.
- Define, provision, and evolve infrastructure as code while managing state and minimizing blast radius.
Requirements
- 5+ years of experience in infrastructure, DevOps, or SRE roles operating production cloud infrastructure.
- Strong Kubernetes experience, including workload deployment, troubleshooting, and work with operators and controllers.
- Strong infrastructure-as-code skills and Terraform experience.
- Production software engineering experience with Python and/or TypeScript, plus Linux and networking fundamentals.
- Experience designing and maintaining CI/CD pipelines and observability platforms such as Datadog or OpenTelemetry.
- Proactive ownership of a broad technical surface area.
Nice to have
- PostgreSQL performance tuning and replica management.
- ML/AI infrastructure experience, including inference services, GPU workloads, model serving, or evaluation pipelines.
- Multi-tenant deployment patterns, per-customer isolation, sandboxed execution environments, or automated reliability systems.
Culture & Benefits
- Competitive base salary with performance-based compensation.
- Equity participation.
- Medical, dental, and vision coverage.
- Flexible paid time off.
- Senior individual-contributor role with broad ownership and significant architectural influence.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
2 дня назад
Senior AI DevOps Developer (AI)
150 000 - 206 000$
22 часа назад
Senior Software Engineer - Platform Engineer (Fintech)
150 000 - 200 000$
Inworld AI
5 дней назад
Staff / Principal Platform Engineer (AI)
23 333 - 29 167$
5 дней назад
Senior DevOps Engineer (AI)
14 167 - 15 833$
18 часов назад
Platform Engineer II (Cloud Infrastructure)
115 000 - 130 000$
7 дней назад
Senior Site Reliability Engineer (Fintech)
8 679 - 14 800$