Назад
Company hidden
обновлено 3 дня назад

Data Center Site Manager (AI/HPC)

Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
Malaysia
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Data Center Site Manager (AI/HPC): Leading and managing daily operations of a mission-critical data center facility with an accent on AI/HPC infrastructure performance and reliability. Focus on supervising shift operations engineers, ensuring 24x7 uptime, and overseeing the maintenance of advanced NVIDIA GPU clusters and networking systems.

Location: Must be based in Johor Bahru, Malaysia

Company

hirify.global is a world-leading technology company specializing in Bitcoin mining and AI cloud solutions with a global footprint.

What you will do

  • Lead and manage daily site operations to ensure infrastructure availability and reliability.
  • Supervise a team of six shift Operations Engineers, including scheduling, coaching, and performance management.
  • Act as the primary escalation point for operational incidents and drive root cause analysis.
  • Oversee maintenance and troubleshooting of AI/HPC infrastructure, including NVIDIA GB200/GB300 clusters and high-speed networking.
  • Establish and improve operational procedures, SOPs, and preventive maintenance programs.
  • Collaborate with engineering and vendor teams to support new deployments and infrastructure expansion.

Requirements

  • Must be based in Johor Bahru, Malaysia.
  • Bachelor's degree in Computer Science, Engineering, or related field.
  • Minimum 5 years of experience in data center operations or HPC/AI infrastructure management.
  • Minimum 2 years of experience in team leadership or people management.
  • Strong knowledge of Linux administration, hardware diagnostics, and networking (Ethernet/InfiniBand).
  • Proven ability to manage 24x7 shift operations in a mission-critical environment.

Nice to have

  • Experience with large-scale AI or HPC clusters.
  • Familiarity with NVIDIA GPU architecture, NVLink, and NVSwitch.
  • Experience with structured cabling systems (optical fiber, DAC, AOC).

Culture & Benefits

  • Inclusive environment with an exciting start-up spirit.
  • Opportunities to network with industrial pioneers in the digital asset industry.
  • Focus on personal accountability, autonomy, and fast professional growth.
  • Attractive welfare benefits and developmental opportunities including training and mentoring.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →