Назад
Company hidden
2 дня назад

HPC Linux System Administrator

Формат работы
remote (только USA)/onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
HPC Linux System Administrator (HPC/Linux): Administering and improving an on-premises NASA human-spaceflight HPC cluster with an accent on job scheduling, parallel filesystems, containerized workflows, and CI/CD integration. Focus on configuring and troubleshooting Slurm, PBS/Torque, or LSF, managing Lustre or GPFS, supporting shared-cluster environments, and maintaining reliable infrastructure for scientists and engineers.

Location: Clear Lake, 2101 E NASA Pkwy, Houston, TX 77058, USA. Fully remote work may be considered for candidates with strong, demonstrated HPC experience. U.S. citizenship or U.S. permanent residency and completion of a U.S. government background investigation are required.

Benefits include medical, dental, vision, life and disability insurance, paid time off, and a 401(k).

Company

hirify.global supports NASA under the JETS II contract, including the Flight Sciences Laboratory at Johnson Space Center.

What you will do

  • Administer and improve a 700+ machine, 26,000-core, 10+ petabyte on-premises HPC environment.
  • Configure, monitor, and troubleshoot an HPC job scheduler such as Slurm, PBS/Torque, or LSF.
  • Administer a high-speed parallel filesystem such as Lustre or GPFS.
  • Support containerized HPC workflows and CI/CD pipelines, including Jacamar and run-node integration.
  • Maintain system configuration, monitoring, alerting, and cluster reliability.
  • Work directly with scientists and engineers supporting NASA human spaceflight programs.

Requirements

  • Production experience administering an HPC job scheduler, including queue or partition configuration, accounting, and job-failure troubleshooting.
  • Production experience administering Lustre or GPFS; general NAS, SAN, or NFS administration is not a substitute.
  • At least five years of Linux system administration experience in an on-premises, bare-metal, cluster, or research-computing environment.
  • Experience with containers in HPC, CI/CD workflows, configuration management, and monitoring or alerting systems.
  • A bachelor's degree or equivalent certification in a related field, or equivalent experience.
  • Proof of U.S. citizenship or U.S. permanent residency and eligibility for a U.S. government background investigation are required.

Nice to have

  • Red Hat-based Linux, InfiniBand, xCAT, Warewulf, Ansible, or Foreman experience.
  • SPACK, EasyBuild, Lmod, Environment Modules, MPI, Git, GitLab, or Jacamar experience.
  • Experience with AI tooling in sysadmin workflows, automation, or user-facing HPC jobs.
  • Knowledge of NASA security mechanisms, including security plans, POAMs, ATOs, and risk assessments.

Culture & Benefits

  • Work in a critical computing environment supporting ISS, Orion, SLS, Commercial Crew, Lunar Gateway, and the Human Landing System.
  • Comprehensive medical, dental, and vision coverage.
  • Company-paid life and disability insurance.
  • Paid time off and a 401(k).
  • 9/80 work schedule with every other Friday off when applicable.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →