1 день назад
Engineer (AI Infrastructure)
200 000 - 275 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Engineer (AI Infrastructure) (C/C++/Rust/Linux): Building foundational control-plane software, system services, telemetry, and orchestration primitives for cluster-scale AI inference deployments with an accent on low-level systems programming and hardware–software integration. Focus on profiling and tuning PCIe, memory, networking, kernel, and runtime layers, debugging complex failures, and bringing large-scale compute systems into stable operation.
Location: On-site in San Jose, Santana Row
Salary: $200,000–$275,000 per year plus significant equity
Company
develops co-designed chips, racks, software, and manufacturing systems focused on high-throughput, low-latency AI inference.
What you will do
- Architect and implement low-level control-plane software for system bring-up, configuration, and management of cluster-scale AI compute deployments.
- Build system services that interact with hardware, firmware, and the operating system.
- Develop telemetry, logging, and tracing infrastructure to diagnose failures and improve performance.
- Implement orchestration primitives for devices, nodes, and racks.
- Profile and tune performance across PCIe, memory, networking, kernel, and runtime layers.
- Collaborate with hardware, firmware, kernel, and runtime teams to co-design system interfaces.
Requirements
- Strong proficiency in C, C++, or Rust for low-level systems programming.
- Deep understanding of Linux internals, kernel and user-space boundaries, and system-level debugging.
- Experience with drivers, DMA, interrupts, memory management, or device control paths.
- Strong debugging skills using logs, tracing, and low-level observability tools.
- Strong communication skills and comfort collaborating across hardware and software teams.
- Availability for fully in-person work in San Jose.
Nice to have
- Experience with Kubernetes, Docker, or other data center orchestration technologies.
- Experience with kernel development, device drivers, firmware-adjacent software, or system bring-up.
- Familiarity with PCIe, NUMA, networking, high-speed interconnects, perf, eBPF, or ftrace.
- Background in HPC, AI infrastructure, large-scale compute systems, system test harnesses, or failure-injection frameworks.
Culture & Benefits
- Fully in-person, cross-disciplinary engineering and research environment.
- Medical, dental, and vision coverage with generous premium support.
- $500 monthly credit for waiving medical benefits.
- $2,000 monthly housing subsidy for employees living within walking distance of the office.
- Relocation support for moves to San Jose, plus wellness benefits and daily lunch and dinner.
- Unlimited compute budget subject to ROI justification.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
Staff Software Engineer (Linux Networking)
185 000 - 224 000$
1 день назад
AI Infrastructure Operations Engineer (SiteOps)
1 день назад
Systems Administrator
120 000 - 148 000$
1 день назад
Site Reliability Engineer
220 000 - 300 000$
1 день назад
Senior IT & Security Operations Engineer (Robotics)
The San Francisco Compute Company
6 дней назад
Software Engineer (Systems)
14 167 - 41 667$