Infrastructure Engineer (Storage)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Location: London, England, United Kingdom; New York, New York, United States; San Francisco, California, United States; or Seattle, Washington, United States. Hybrid role based in one of these hubs with a minimum of 2 in-office days per week and occasional team and company offsites.
Annual base salary: $180,000–$220,000 USD, plus discretionary bonus, equity, and benefits.
Company
builds an end-to-end platform for developing, training, and deploying AI systems, combining developer-focused software with large-scale cloud and compute infrastructure.
What you will do
- Operate and scale distributed storage systems, including VAST and S3-compatible object storage such as Ceph.
- Improve storage performance, reliability, and efficiency for high-throughput, low-latency AI/ML and HPC workloads.
- Build Python-based automation for provisioning, monitoring, lifecycle management, maintenance, and scaling.
- Manage Linux-based production systems and bare-metal infrastructure.
- Troubleshoot complex issues across storage, operating systems, hardware, and networking.
- Partner with infrastructure, network, platform, and data center teams on deployments, hardware operations, capacity planning, and scaling strategies.
Requirements
- 5+ years of experience in infrastructure engineering, systems engineering, or related roles.
- Hands-on experience operating distributed storage systems such as VAST, Ceph, or similar platforms.
- Strong production Linux experience and proficiency in Python or a similar programming language for automation.
- Experience with bare-metal infrastructure, hardware-oriented systems, and storage networking protocols such as NFS.
- Ability to debug issues across system boundaries and experience with capacity planning, monitoring, and performance tuning.
- Ability to work from one of the listed London or U.S. hubs and attend the office at least two days per week; visa sponsorship is not available.
Nice to have
- Production experience with VAST and S3-compatible object storage at scale.
- Data center operations and physical hardware experience.
- Experience with AI/ML, HPC, GPU-based workloads, or large-scale compute clusters.
- Background in high-performance or low-latency distributed systems.
- Familiarity with RDMA or GPU Direct Storage.
Culture & Benefits
- Ownership-oriented environment focused on urgency, open communication, continuous improvement, and long-term scalable systems.
- Medical, dental, and vision coverage for employees and eligible dependents.
- Equity, U.S. 401(k) matching, and U.K. pension contributions.
- Unlimited PTO, company holidays, floating holidays, and a two-week winter break.
- Paid parental and family leave, professional development allowance, wellness and work-from-home stipends, and a sabbatical after four years.
- Complimentary meals at office hubs and flexible schedules within the hybrid work model.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →