1 месяц назад
Infrastructure Operations Deputy Manager (AI Infrastructure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Infrastructure Operations Deputy Manager (AI Infrastructure): Managing devices, systems, and infrastructure for an AI data center with an accent on HPC and GPU operations, 24/7 reliability, and power, cooling, and environmental controls. Focus on leading infrastructure teams, maintaining SLA and KPI performance, troubleshooting complex technical issues, and coordinating vendors, inventory, and continuous operational improvements.
Location: Norway; based on-site at the data center. Occasional travel to other locations and participation in an on-call rotation are required.
Company
Nscale provides a cost-effective, high-performance GPU cloud infrastructure platform for AI startups and enterprise customers.
What you will do
- Take overall responsibility for devices, systems, and infrastructure across the AI data center.
- Ensure 24/7 operational reliability and performance for high-performance AI workloads.
- Manage power, cooling, environmental conditions, installations, configuration, and maintenance of HPC and GPU systems.
- Lead, mentor, train, and schedule engineers, technicians, and support staff, including on-call coverage.
- Act as the primary client point of contact for SLA and KPI reporting, while coordinating vendors, contractors, repairs, upgrades, and procurement.
- Monitor performance and resource utilization, troubleshoot technical issues, manage spare-parts inventory, and implement reliability, scalability, and cost improvements.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related field.
- At least 5 years of experience managing data centers, particularly in HPC and GPU environments.
- Proven experience leading and developing technical teams.
- Strong expertise in HPC, GPU deployment, hardware and software solutions, network architecture, and performance optimization.
- Strong understanding of data-center power, cooling, and environmental systems.
- Excellent client-facing communication and reporting skills, plus experience with inventory and spare-parts management.
Nice to have
- Data center management certifications such as CDCP or CDCS.
- Hands-on experience with NVIDIA GPUs, CUDA, and AI frameworks.
- Familiarity with hybrid cloud and HPC environments.
Culture & Benefits
- Opportunity to shape operating standards for a next-generation AI cloud platform.
- Work on complex infrastructure challenges with significant ownership and operational impact.
- Focus on scaling high-performance and sustainable data-center operations.
- Continuous learning is encouraged to keep pace with evolving technologies and standards.
- Inclusive and equitable workplace with support for individual accessibility needs.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →