5 часов назад
AI Infrastructure Solutions Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
AI Infrastructure Solutions Engineer (NVIDIA AI Enterprise, Milvus, Storage): Deploying and customizing production-scale AI, RAG, and HPC solutions integrating NVIDIA AI Enterprise services, vector databases, high-performance storage, and networking fabrics with an accent on customer environments and repeatable deployment workflows. Focus on automating infrastructure operations, tuning distributed storage and networking, and optimizing AI applications across on-premises, hybrid, and cloud-adjacent environments.
Location: Remote in Spain, with onsite deployment work required in some cases
Company
provides enterprise AI and high-performance storage solutions based on Infinia and EXAScaler.
What you will do
- Deploy and customize AI solutions for strategic enterprise customers.
- Integrate NVIDIA AI Enterprise services, vector databases, RAG and agentic workflows, storage, and networking.
- Serve as the primary technical contact for assigned customers and interpret deployment designs with Pre-sales.
- Develop automation, scripts, monitoring tools, documentation, and repeatable deployment workflows.
- Support AI/ML, data-intensive, and HPC workloads across on-premises, hybrid, and cloud-adjacent environments.
- Collaborate with Engineering, Product Management, Support, Field Services, vendors, and partners to optimize customer applications.
Requirements
- 5+ years of experience in a senior technical role deploying complex, customer-facing production systems.
- Experience administering Lustre or similar parallel file systems in large-scale environments.
- Strong Linux systems knowledge, including performance tuning and troubleshooting.
- Understanding of distributed storage architectures, networking fundamentals, and data movement at scale.
- Ability to work directly with customers and collaborate across cross-functional technical teams.
- Occasional physical hardware setup tasks may be required with appropriate tools and support.
Nice to have
- Experience with IBM Spectrum Scale, StorNext, object storage, S3-compatible systems, NAS protocols, and data transfer tools.
- Automation development and debugging with shell scripting, Python, or Bash.
- Knowledge of InfiniBand, Ethernet, TCP/IP, routing, and network diagnostics tools.
- Knowledge of NVIDIA AI Enterprise services, including NIMs, NeMo, Triton, TensorRT, GPU Operator, and licensing, plus Milvus.
- Experience with Kubernetes, SLURM, BCM, authentication systems, GPU environments, and large-scale data pipelines.
Culture & Benefits
- Work on production-scale AI and HPC systems supporting advanced innovation.
- Collaborate with engineers specializing in storage, networking, and distributed systems.
- Influence product direction through direct customer engagement.
- Opportunity to grow toward senior technical leadership, architecture, or product-facing roles.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Latitude
1 день назад
Senior Operations Analyst (Infrastructure)
6 дней назад
Lead Middleware Systems Administrator (Middleware)
3 часа назад
HQ - Junior IT Helpdesk Engineer
Sateliot
2 дня назад
NOC Operator (Satellite Networks)
2 дня назад
IT Support Engineer L2 (SaaS Administration)
7 дней назад