6 дней назад
Technical Lead (Distributed Storage)
270 000 - 330 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Technical Lead (Distributed Storage): Building and operating the distributed object storage system underpinning container images, volumes, and checkpoints for a high-performance serverless AI platform with an accent on petabyte-scale data, multi-cloud replication, and low-level storage infrastructure. Focus on leading engineers, designing active-active replication and large-scale garbage collection, and balancing durability, latency, throughput, and cost across a growing fleet.
Location: On-site in San Francisco or New York, United States
Salary: $270,000–$330,000 per year
Company
is building an infrastructure layer for AI workloads through a high-performance serverless platform.
What you will do
- Lead a team of 3–8 engineers while remaining hands-on across the storage stack.
- Design, build, and maintain the distributed object storage system supporting container images, volumes, and checkpoints.
- Set technical direction for storage primitives used by filesystem, training, and sandbox teams.
- Own the roadmap for petabyte-scale garbage collection, active-active replication, rate limiting, colocated GPU storage, and tiered writes.
- Guide observability, automation, capacity planning, and on-call practices for systems operating across multiple datacenters and cloud providers.
Requirements
- 7+ years of experience writing high-quality production code.
- 3+ years of direct people management experience.
- Experience building and operating high-performance distributed storage or caching systems at large scale.
- Strong knowledge of object storage such as S3, CDNs, Linux kernel foundations, filesystems, page cache, and containers.
- Experience with replication, content addressing, consistency models, petabyte-scale datasets, data migration, cost planning, and capacity planning.
- Willingness to participate in the on-call rotation and respond to production incidents.
Nice to have
- Experience with data engineering at petabyte scale.
- Prior experience with Rust.
Culture & Benefits
- Work on infrastructure supporting high-throughput AI workloads.
- Collaborate with engineers who have created open-source projects, conducted academic research, and led engineering and product organizations.
- Operate systems involving hundreds of petabytes of data, multiple cloud object stores, CDNs, local NVMe caches, and peer-to-peer datacenter sharing.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
Principal Software Engineer (Distributed Systems)
220 000 - 250 000$
Snowflake
10 дней назад
Engineering Manager, Cloud Efficiency
236 000 - 339 200$
13 дней назад
Backend Engineer (LiveOps)
142 000 - 265 000$
11 дней назад
Forward Deployed Engineer (AI)
10 дней назад
Lead Forward Deployed Engineer
7 дней назад