Назад
Company hidden
1 день назад

Infrastructure Engineer (Storage)

180 000 - 220 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK/US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Infrastructure Engineer (Storage) (AI/HPC): Building and operating distributed storage systems for large-scale AI/ML training, inference, and HPC workloads with an accent on high-throughput data access, low latency, and reliability across bare-metal infrastructure. Focus on automating storage lifecycle management, debugging issues across hardware and software boundaries, and scaling VAST and S3-compatible systems such as Ceph.

Location: London, England, United Kingdom; New York, New York, United States; San Francisco, California, United States; or Seattle, Washington, United States. Hybrid role based in one of these hubs with a minimum of 2 in-office days per week and occasional team and company offsites.

Annual base salary: $180,000–$220,000 USD, plus discretionary bonus, equity, and benefits.

Company

hirify.global builds an end-to-end platform for developing, training, and deploying AI systems, combining developer-focused software with large-scale cloud and compute infrastructure.

What you will do

  • Operate and scale distributed storage systems, including VAST and S3-compatible object storage such as Ceph.
  • Improve storage performance, reliability, and efficiency for high-throughput, low-latency AI/ML and HPC workloads.
  • Build Python-based automation for provisioning, monitoring, lifecycle management, maintenance, and scaling.
  • Manage Linux-based production systems and bare-metal infrastructure.
  • Troubleshoot complex issues across storage, operating systems, hardware, and networking.
  • Partner with infrastructure, network, platform, and data center teams on deployments, hardware operations, capacity planning, and scaling strategies.

Requirements

  • 5+ years of experience in infrastructure engineering, systems engineering, or related roles.
  • Hands-on experience operating distributed storage systems such as VAST, Ceph, or similar platforms.
  • Strong production Linux experience and proficiency in Python or a similar programming language for automation.
  • Experience with bare-metal infrastructure, hardware-oriented systems, and storage networking protocols such as NFS.
  • Ability to debug issues across system boundaries and experience with capacity planning, monitoring, and performance tuning.
  • Ability to work from one of the listed London or U.S. hubs and attend the office at least two days per week; visa sponsorship is not available.

Nice to have

  • Production experience with VAST and S3-compatible object storage at scale.
  • Data center operations and physical hardware experience.
  • Experience with AI/ML, HPC, GPU-based workloads, or large-scale compute clusters.
  • Background in high-performance or low-latency distributed systems.
  • Familiarity with RDMA or GPU Direct Storage.

Culture & Benefits

  • Ownership-oriented environment focused on urgency, open communication, continuous improvement, and long-term scalable systems.
  • Medical, dental, and vision coverage for employees and eligible dependents.
  • Equity, U.S. 401(k) matching, and U.K. pension contributions.
  • Unlimited PTO, company holidays, floating holidays, and a two-week winter break.
  • Paid parental and family leave, professional development allowance, wellness and work-from-home stipends, and a sabbatical after four years.
  • Complimentary meals at office hubs and flexible schedules within the hybrid work model.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →