4 дня назад
AI System Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
AI System Engineer (AI agent systems and accelerated computing): Building and maintaining Wingman and an agentic AI platform, including agent capabilities, secure execution environments, workload orchestration, and software integrations with an accent on computer vision models, LLM services, and inference infrastructure. Focus on validating platform correctness on real hardware, investigating failures across model and compute stacks, and ensuring reliable execution, security, and recovery.
Location: Flexible arrangement with hybrid or on-site work available from offices in Florence, Milan, Amsterdam, Eindhoven, Leuven, Paris, Zurich, and Bristol; fully remote work is available from any European country, including the UK.
Company
is a deep-tech company developing the Metis AI Platform and agentic AI technologies for accelerated computing and AI workloads.
What you will do
- Build reliable agent systems, intelligent tools, and application integrations for Wingman and the agentic AI platform.
- Develop secure, scalable execution environments for demanding AI workloads, including remote workload execution and accelerated computing access.
- Deploy and optimize computer vision models, LLM services, and inference infrastructure to validate platform capabilities.
- Own capabilities from implementation and hardware validation through production operation and ongoing reliability.
- Test functionality, security boundaries, hardware behavior, failure recovery, and workload performance.
- Write clear documentation, reproducible setup instructions, and operational runbooks while collaborating across technical areas.
Requirements
- Practical experience building with AI agents, model APIs, and tools, including tool calling, structured outputs, streaming, or MCP.
- Understanding of agent state management, application interaction, failure recovery, and effective use of coding agents.
- Experience with Linux, processes, resource allocation, secure permissions, credentials, workload isolation, and failure diagnosis across software and compute stacks.
- Hands-on experience deploying computer vision models and working with LLM services and inference workloads on real hardware.
- Ability to assess accuracy, latency, throughput, and resource usage and investigate failures across model code, inference SDKs, and execution environments.
- Demonstrated technical ability and judgment; specific degrees or certifications are not required.
Nice to have
- Python, Rust, TypeScript, React, Tauri, APIs, databases, and native integrations.
- Workload orchestration, scheduling, quotas, concurrency, reproducible environments, and SDK or driver compatibility management.
- Authentication, authorization, least privilege, sandboxing, network security, and protection of user files and data.
- Google Cloud or AWS, containers, infrastructure as code, CI/CD, monitoring, and safe deployment or recovery practices.
- PyTorch, ONNX, image and video processing, model evaluation, retrieval, fine-tuning, multimodal applications, distributed inference, or desktop packaging.
Culture & Benefits
- Flexible choice between office-based, hybrid, and remote work within Europe.
- International organization with an open culture focused on creativity, innovation, collaboration, and ownership.
- Attractive compensation package with a pension plan and extensive employee insurances.
- Option to receive company shares.
- Inclusive equal-opportunity environment welcoming applicants from all backgrounds.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →