Назад
Company hidden
4 дня назад

Staff Engineer (LustreFS)

200 000 - 250 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff Engineer (LustreFS): Designing, debugging, and optimizing distributed storage features and fixes across LustreFS subsystems with an accent on Linux systems engineering, performance, reliability, and operational excellence. Focus on investigating customer and scale-related failures, resolving issues across client, server, network, and storage layers, and using AI-assisted workflows to accelerate triage and code understanding.

Location: Santa Clara Office, hybrid

Salary: $200K–$250K annually

Company

hirify.global's Engineering team develops and supports LustreFS-based distributed storage systems.

What you will do

  • Design, develop, and debug LustreFS features, fixes, and enhancements across llite, MDS/MDT, OSS/OST, LDLM, and LNet.
  • Investigate customer and scale-related defects, lead root-cause analysis, and implement maintainable fixes.
  • Improve performance, failure handling, reliability, observability, and serviceability for large-scale Lustre deployments.
  • Collaborate with QE, support, engineering, and release teams to reproduce issues, improve diagnostics, and expand coverage.
  • Review designs and code, validate fixes through targeted tests, failure injection, log analysis, and performance characterization.
  • Use AI-assisted tools for triage, debugging, code comprehension, documentation, and early-stage design work.

Requirements

  • 10+ years of experience in systems software, distributed systems, storage, Linux kernel, or filesystem engineering.
  • Strong experience with LustreFS development, support, or performance engineering, including depth in at least one major subsystem.
  • Strong C programming and Linux systems debugging skills.
  • Working knowledge of Linux kernel internals, filesystem semantics, networking, and performance analysis.
  • Experience with LNet or high-performance transports such as RDMA, InfiniBand, RoCE, or TCP-based storage networking.
  • Ability to debug issues spanning client, server, network, and backend storage layers.

Nice to have

  • Experience in HPC, AI infrastructure, or large-scale parallel storage environments.
  • Familiarity with ZFS, ldiskfs, NVMe-backed storage, and observability or performance tooling.
  • Experience creating test plans, reproducer frameworks, runbooks, or diagnostic automation.
  • Experience mentoring junior engineers or leading focused technical efforts within a subsystem.

Culture & Benefits

  • Hands-on collaboration across engineering, QE, support, and release functions.
  • Emphasis on documentation, code review quality, subsystem knowledge sharing, and engineering redundancy.
  • Focus on durable LustreFS ownership, operational readiness, and AI-enabled engineering execution.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →