4 дня назад
Senior MLOps / ML Platform Engineer (AdTech)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior MLOps / ML Platform Engineer (AdTech): Building a production-grade ML platform for a high-load programmatic advertising ecosystem with an accent on orchestration, model lifecycle automation, and observability. Focus on designing reproducible multi-tenant workflows, real-time optimization infrastructure, and monitoring for drift, train/serve skew, and operational reliability.
Location: Remote across Latin America and Europe, including Argentina, Brazil, Colombia, Mexico, Poland, Romania, and Ukraine
Company
Technology and engineering services provider building large-scale ML infrastructure for a programmatic advertising customer.
What you will do
- Build and maintain ML training orchestration pipelines with hourly, daily, and weekly schedules, including retries, backfills, and idempotent execution.
- Design model registry workflows covering versioning, lineage, evaluation gates, and promotion processes.
- Develop isolated per-advertiser model environments with namespace and configuration separation.
- Build refresh and publishing pipelines for serving infrastructure, including shadow mode and champion/challenger deployments.
- Implement monitoring and alerting for feature drift, prediction drift, train/serve skew, calibration decay, and operational costs.
- Collaborate with DevOps and SRE engineers on CI/CD, infrastructure automation, documentation, and platform handover.
Requirements
- 5+ years of experience in MLOps, ML platform engineering, or infrastructure engineering for production ML systems.
- Expert Python and MLOps skills, with experience building platform tooling and automation.
- Strong hands-on experience with Kubernetes, Docker, CI/CD for ML workloads, and Linux environments.
- Production experience with MLflow, Kubeflow, Airflow, Argo Workflows, Vertex Pipelines, Vertex AI, or similar platforms.
- Experience with multi-tenant ML systems, model lifecycle management, observability, infrastructure as code, and preferably GCP.
- Upper-Intermediate English level or higher required.
Nice to have
- Experience with feature stores, large-scale batch scoring, experiment tracking, evaluation gates, or data versioning tools.
- Experience with on-premises Kubernetes, bare-metal Linux, low-latency serving databases, GPU scheduling, or training cost optimization.
- Knowledge of SOC 2, ISO 27001, or GDPR-related compliance requirements.
Culture & Benefits
- Remote work with a flexible schedule.
- Health and legal support.
- Continuous education and an active professional community.
- Professional and community activities, including sports.
- Office access is available if preferred.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →