21 день назад
Platform Engineer (GPU/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Platform Engineer (GPU/Kubernetes) (Robotics Simulation): Building and operating GPU-backed infrastructure for photorealistic digital twins, Gaussian Splatting workflows, and browser-based Omniverse and Isaac Sim streaming with an accent on Kubernetes, reproducible deployments, and multi-tenant operations. Focus on designing scheduling and autoscaling for interactive and compute-intensive workloads, productizing deployments for customer environments, and establishing observability, security, and self-service platform foundations.
Location: Dresden, Germany; onsite
Company
develops NOVA, a platform that makes industrial robot programming accessible through simulation and digital twins.
What you will do
- Build and operate GPU infrastructure for simulation workloads, including NVIDIA GPU Operator, device plugins, driver lifecycle, MIG, and time slicing.
- Integrate simulation workloads with shared Kubernetes and cloud infrastructure using Infrastructure as Code, GitOps, and reproducible deployments.
- Own CI/CD and container image lifecycles for CUDA- and NGC-based workloads.
- Design scheduling and autoscaling for interactive streaming sessions and compute-intensive Gaussian Splatting jobs.
- Operate and extend Omniverse and Isaac Sim streaming, including session lifecycle, tenant isolation, observability, and security.
- Productize the platform for internal teams, partners, and customers through self-service workflows, versioned releases, upgrade paths, diagnostics, and documentation.
Requirements
- Experience operating production systems in platform engineering, infrastructure, SRE, DevOps, or systems software.
- Practical Kubernetes experience with resource management, workload scheduling, and distributed-system debugging.
- Experience with Infrastructure as Code, GitOps, CI/CD, Linux, containers, and networking.
- Good Python skills for automation, services, and data-processing workflows.
- Experience with GPU workloads or a strong foundation and interest in drivers, CUDA, and GPU scheduling.
- Very good English is required; German is a plus.
Nice to have
- Experience with NVIDIA Omniverse, Isaac Sim, Kit SDK, OpenUSD, Gaussian Splatting, NeRF, or photogrammetry.
- Knowledge of WebRTC, pixel streaming, robotics, ROS 2, or sim-to-real workflows.
- Experience with multi-tenant platforms, air-gapped deployments, or customer-controlled environments.
- Python, Rust, or Go experience for infrastructure tooling and services.
Culture & Benefits
- Work on a deep-tech robotics product in an international, cross-functional environment.
- Access to knowledge sharing, training, and mental health opportunities.
- Up to 3 days of volunteer time off.
- Free lunch, snacks, drinks, and team events.
- 30 days of annual vacation plus additional rest during the Christmas season.
- Full pay when caring for a sick child.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →