3 дня назад
Senior Platform Engineer (AI/GPU)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Platform Engineer (AI/GPU): Building and operating a GCP-based GPU platform and self-service Kubernetes infrastructure for AI model training and serving with an accent on infrastructure-as-code, distributed systems, and platform reliability. Focus on designing GPU workloads, enabling data science teams through job scheduling and cost visibility, and establishing automated deployment paths for production AI services.
Location: Office-based in Paris, France; Montreal, Canada; or Bucharest, Romania
Company
is a global video game company creating original gaming experiences across international studios.
What you will do
- Design, build, and operate a GPU platform on GCP for AI workloads using infrastructure-as-code.
- Deliver self-service compute for data science teams, including Ray on Kubernetes, job submission, queuing, quotas, and cost visibility.
- Establish golden paths that enable product teams to deploy services without repeating operational work.
- Support the deployment, hosting, and production operation of machine learning models.
- Control infrastructure costs and improve system performance, reliability, and efficiency.
- Collaborate with data science teams from proof of concept through production deployment and automate internal workflows with AI-powered systems.
Requirements
- Significant experience as a Platform, DevOps, or MLOps Engineer with strong software engineering fundamentals and distributed systems knowledge.
- Hands-on experience with Kubernetes, Terraform, GitOps, CRDs and operators, scheduling, autoscaling, node pool design, state management, ArgoCD, and Helm.
- Solid experience with GCP.
- Ability to work autonomously with a platform-as-a-product mindset, document solutions, measure adoption, and support data scientists as platform users.
- Fluent English, written and spoken, is required.
Nice to have
- Experience running or optimizing GPU workloads for machine learning training or inference in production.
- Experience with distributed job scheduling such as KubeRay, Slurm, or Kubeflow.
- Experience with cross-cloud networking between AWS and GCP.
- Familiarity with MLOps tools such as MLflow, Weights & Biases, and model registries.
- Real-world AWS experience.
Culture & Benefits
- Access to an internal e-learning platform from the first day.
- Access to a game library with titles, competitor games, consoles, and board games.
- Works council discounts for entertainment, fitness, cultural activities, and other services.
- Career and development planning with a manager after one year.
- Clubs, gym classes, bikes, organized sports weekends, and recreational activities.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Engineering Manager, DevEx (Kubernetes)
4 дня назад
Senior Cloud Infrastructure Engineer (AWS/Kubernetes)
4 дня назад
Senior Platform Developer (AWS, Cloud Platform)
7 дней назад
Platform Engineer (AI)
140 000 - 230 000CAD
Robinhood
10 дней назад
Senior Staff Software Developer, Developer Infrastructure (AI)
234 000 - 275 000CAD
9 дней назад