4 дня назад
Staff Engineer (LustreFS)
200 000 - 250 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Engineer (LustreFS): Designing, debugging, and optimizing distributed storage features and fixes across LustreFS subsystems with an accent on Linux systems engineering, performance, reliability, and operational excellence. Focus on investigating customer and scale-related failures, resolving issues across client, server, network, and storage layers, and using AI-assisted workflows to accelerate triage and code understanding.
Location: Santa Clara Office, hybrid
Salary: $200K–$250K annually
Company
's Engineering team develops and supports LustreFS-based distributed storage systems.
What you will do
- Design, develop, and debug LustreFS features, fixes, and enhancements across llite, MDS/MDT, OSS/OST, LDLM, and LNet.
- Investigate customer and scale-related defects, lead root-cause analysis, and implement maintainable fixes.
- Improve performance, failure handling, reliability, observability, and serviceability for large-scale Lustre deployments.
- Collaborate with QE, support, engineering, and release teams to reproduce issues, improve diagnostics, and expand coverage.
- Review designs and code, validate fixes through targeted tests, failure injection, log analysis, and performance characterization.
- Use AI-assisted tools for triage, debugging, code comprehension, documentation, and early-stage design work.
Requirements
- 10+ years of experience in systems software, distributed systems, storage, Linux kernel, or filesystem engineering.
- Strong experience with LustreFS development, support, or performance engineering, including depth in at least one major subsystem.
- Strong C programming and Linux systems debugging skills.
- Working knowledge of Linux kernel internals, filesystem semantics, networking, and performance analysis.
- Experience with LNet or high-performance transports such as RDMA, InfiniBand, RoCE, or TCP-based storage networking.
- Ability to debug issues spanning client, server, network, and backend storage layers.
Nice to have
- Experience in HPC, AI infrastructure, or large-scale parallel storage environments.
- Familiarity with ZFS, ldiskfs, NVMe-backed storage, and observability or performance tooling.
- Experience creating test plans, reproducer frameworks, runbooks, or diagnostic automation.
- Experience mentoring junior engineers or leading focused technical efforts within a subsystem.
Culture & Benefits
- Hands-on collaboration across engineering, QE, support, and release functions.
- Emphasis on documentation, code review quality, subsystem knowledge sharing, and engineering redundancy.
- Focus on durable LustreFS ownership, operational readiness, and AI-enabled engineering execution.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
Senior Systems Engineer (AI)
250 000 - 400 000$
Nscale
5 дней назад
Staff HPC Systems Software Engineer (AI)
225 000 - 275 000$
11 дней назад
Systems and Solutions Engineer (Networking)
122 440 - 172 860$
5 дней назад
Data Center Engineer
98 000 - 165 000$
Anthropic
8 дней назад
Staff Engineer, Datacenter Server Lifecycle
320 000 - 405 000$
Microsoft AI
10 дней назад
Member of Technical Staff - AI Networking
119 800 - 234 700$