Назад
Company hidden
21 час назад

Lab Systems Administrator

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Lab Systems Administrator (Linux/Storage): Operating and maintaining a technical research lab environment covering servers, storage systems, networking equipment, and workstations with an accent on availability, performance, security, and compliance. Focus on provisioning experimental infrastructure, automating administration with Bash and Python, troubleshooting hardware and software, and managing backups, monitoring, and system capacity.

Location: Columbia Office; On-site

Company

hirify.global supports a technical and research lab environment with servers, storage systems, networking equipment, and lab workstations.

What you will do

  • Install, configure, and maintain lab servers, storage arrays, network interfaces, switches, operating system images, drivers, and monitoring tools.
  • Manage user accounts, permissions, system health, performance, capacity, backups, disaster recovery, and data retention.
  • Troubleshoot hardware and software issues and coordinate vendor support and RMAs.
  • Support researchers and engineers by provisioning compute nodes, storage volumes, and network configurations for experiments, benchmarks, and testing.
  • Maintain network diagrams, asset inventories, configuration baselines, and standard operating procedures.
  • Manage security controls, licensing, warranties, hardware lifecycle, and shared lab resource allocation.

Requirements

  • Strong Linux administration experience with Ubuntu, RHEL, or CentOS.
  • Knowledge of TCP/IP, VLANs, bonding/LACP, and basic network troubleshooting.
  • Experience with SAN/NAS storage systems; parallel filesystems such as Lustre or GPFS are advantageous.
  • Scripting ability in Bash and Python for automation.
  • Experience with monitoring tools such as Prometheus/Grafana, Nagios, or Zabbix.
  • Understanding of CPUs, NUMA architecture, PCIe, NICs, and RDMA, plus strong documentation and communication skills.

Nice to have

  • Experience with virtualization or containerization using KVM or Docker.
  • Experience with HPC or AI infrastructure, including InfiniBand, RDMA, or GPU clusters.
  • Experience with configuration management tools such as Ansible or Puppet.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →