Назад
Company hidden
6 дней назад

Technical Lead (Distributed Storage)

270 000 - 330 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Technical Lead (Distributed Storage): Building and operating the distributed object storage system underpinning container images, volumes, and checkpoints for a high-performance serverless AI platform with an accent on petabyte-scale data, multi-cloud replication, and low-level storage infrastructure. Focus on leading engineers, designing active-active replication and large-scale garbage collection, and balancing durability, latency, throughput, and cost across a growing fleet.

Location: On-site in San Francisco or New York, United States

Salary: $270,000–$330,000 per year

Company

hirify.global is building an infrastructure layer for AI workloads through a high-performance serverless platform.

What you will do

  • Lead a team of 3–8 engineers while remaining hands-on across the storage stack.
  • Design, build, and maintain the distributed object storage system supporting container images, volumes, and checkpoints.
  • Set technical direction for storage primitives used by filesystem, training, and sandbox teams.
  • Own the roadmap for petabyte-scale garbage collection, active-active replication, rate limiting, colocated GPU storage, and tiered writes.
  • Guide observability, automation, capacity planning, and on-call practices for systems operating across multiple datacenters and cloud providers.

Requirements

  • 7+ years of experience writing high-quality production code.
  • 3+ years of direct people management experience.
  • Experience building and operating high-performance distributed storage or caching systems at large scale.
  • Strong knowledge of object storage such as S3, CDNs, Linux kernel foundations, filesystems, page cache, and containers.
  • Experience with replication, content addressing, consistency models, petabyte-scale datasets, data migration, cost planning, and capacity planning.
  • Willingness to participate in the on-call rotation and respond to production incidents.

Nice to have

  • Experience with data engineering at petabyte scale.
  • Prior experience with Rust.

Culture & Benefits

  • Work on infrastructure supporting high-throughput AI workloads.
  • Collaborate with engineers who have created open-source projects, conducted academic research, and led engineering and product organizations.
  • Operate systems involving hundreds of petabytes of data, multiple cloud object stores, CDNs, local NVMe caches, and peer-to-peer datacenter sharing.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →