Назад
Company hidden
3 дня назад

Senior Storage Engineer (AI)

Формат работы
remote (только Europe)/hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Europe
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Storage Engineer (Ceph/AI): Designing and operating petabyte-scale Ceph clusters for object, block, and file storage that support demanding GPU workloads with an accent on storage architecture, reliability, and managed object storage. Focus on scaling distributed storage to hundreds of petabytes, troubleshooting complex performance issues, automating operations, and leading technical direction across the storage area.

Location: EU / Remote; hybrid setup with locations in Helsinki and London

Company

hirify.global is building a European cloud computing platform for deploying and scaling applications and workloads on infrastructure powered by 100% renewable energy.

What you will do

  • Lead the design, deployment, and operation of large-scale Ceph clusters for production workloads.
  • Set the technical direction for storage architecture, standards, operational practices, and roadmap.
  • Bring a managed Object Storage product to market, including access keys, bucket and object management, SLOs, monitoring, and alerting.
  • Scale and maintain storage systems from petabytes to tens or hundreds of petabytes across CephFS, RBD, and RADOS Gateway.
  • Mentor engineers through reviews, pairing, and shared on-call, while improving observability, automation, and operational tooling.
  • Integrate storage with GPU clusters, lead capacity planning and upgrades, troubleshoot incidents, and drive post-mortem improvements.

Requirements

  • Deep expertise in Ceph deployment, cluster architecture, operations, troubleshooting, and performance tuning.
  • Experience operating Ceph at multi-petabyte scale in production.
  • Technical leadership experience owning a storage or infrastructure area end to end.
  • Experience mentoring engineers and operating mission-critical production infrastructure, including on-call.
  • Strong knowledge of Linux systems and internals, networking, distributed systems debugging, and infrastructure automation.
  • Strong collaboration and communication skills for cross-team and customer-facing work.

Nice to have

  • Experience operating RADOS Gateway as a customer-facing managed Object Storage product.
  • Production experience with CephFS and knowledge of Ansible.
  • Familiarity with Prometheus, Grafana, Loki, or similar observability tools.
  • Experience operating storage at a cloud provider or at 10–100 PB scale.
  • Familiarity with GPU and AI workload storage patterns.

Culture & Benefits

  • Full-time, permanent employment.
  • Cash and equity compensation.
  • Healthcare, lunch, wellbeing, and other fringe benefits.
  • Work alongside engineers, researchers, and partners across the global AI ecosystem.
  • Low hierarchy and a collaborative, international environment.

Hiring process

  • Applications are submitted through the Careers page; email applications are not accepted.
  • Hiring proceeds without an artificial deadline and moves forward once a suitable candidate is found.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →