Назад
Company hidden
5 часов назад

Data Center Operations Engineer (Linux/GPU)

9 154 - 17 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Data Center Operations Engineer (Linux/GPU): Supporting and deploying critical data center infrastructure with an accent on Linux systems, GPU server deployments, and InfiniBand networking. Focus on cluster bring-up, hardware troubleshooting, and maintaining scalable compute environments for high-performance workloads.

Location: On-site in San Jose, CA

Salary: $57.21 – $106.25 per hour

Company

hirify.global is a leading provider of electronic design automation (EDA) software and hardware tools.

What you will do

  • Deploy and maintain critical data center infrastructure, focusing on Linux-based systems and GPU server environments.
  • Perform cluster bring-up, including InfiniBand fabric configuration, switch management, and subnet troubleshooting.
  • Handle hands-on hardware installation, including rack and stack, cabling, and component replacement (CPUs, RAM, NICs).
  • Conduct daily health checks and utilize monitoring frameworks to proactively identify and mitigate system risks.
  • Coordinate with vendors for hardware diagnostics and warranty services.
  • Develop and maintain technical documentation, operational runbooks, and incident reports.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, IT, or equivalent practical experience.
  • Strong hands-on experience with Linux system administration and performance validation.
  • Proficiency in Linux command-line tools and shell scripting (Bash).
  • Experience with GPU server validation and end-to-end testing in InfiniBand-based clusters.
  • Solid understanding of networking fundamentals, including TCP/IP and the OSI model.
  • Must be based in San Jose and able to lift equipment weighing 50+ lbs in a data center environment.

Nice to have

  • Experience supporting HPC, AI, or large-scale GPU environments.
  • Exposure to data center monitoring tools.
  • Experience with large-scale data center buildouts or refresh programs.

Culture & Benefits

  • Medical, dental, and vision insurance plans.
  • 401(k) plan with employer match.
  • Paid vacation and paid holidays.
  • Employee stock purchase plan.
  • Professional environment focused on developing leaders and innovators.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →