1 день назад
Infrastructure Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Infrastructure Engineer (AI) (Linux/networking): Maintaining, provisioning, and troubleshooting high-performance on-premise servers and networking infrastructure with an accent on Linux administration, x86 hardware, and 100G/400G connectivity. Focus on installing and diagnosing server hardware, configuring firmware and network layers, automating provisioning with Bash and Python, and minimizing downtime during component replacement.
Location: On-site at the Sunnyvale headquarters office, with daily on-site work required
Company
builds specialized AI hardware and infrastructure designed to deliver high-performance model training and inference.
What you will do
- Install, rack, cable, maintain, troubleshoot, and replace blade servers and components including CPUs, DIMMs, NICs, and storage devices.
- Connect servers to 100G/400G networks, verify optics and DACs, and diagnose physical and operating-system-level connectivity issues.
- Configure BIOS, firmware, IPMI/iDRAC/iLO, Linux operating systems, hostnames, IP addresses, routing, and NFS mount points.
- Use Linux tools and Bash or Python scripts to isolate issues and automate provisioning, diagnostics, and routine infrastructure tasks.
- Follow provisioning playbooks, maintain accurate asset and change records, and document procedures.
- Collaborate with network, systems, and storage teams and coordinate vendor RMAs while minimizing downtime.
Requirements
- 3–5+ years of experience in data center, laboratory, or infrastructure engineering.
- Proficiency in Linux system administration and network configuration.
- Strong hands-on knowledge of x86 server hardware and enterprise networking.
- Experience with BIOS configuration, firmware updates, remote management tools, high-speed NICs, and optical links.
- Experience with VLANs, static routing, and diagnosing Layer 1–3 issues.
- Ability to work on-site daily and lift or move server hardware.
Nice to have
- Experience with PXE, NFS, RAID controllers, monitoring tools, or Ansible.
- Prior experience in a laboratory or R&D hardware/software environment.
Culture & Benefits
- Work with a specialized AI platform and high-performance AI supercomputing infrastructure.
- Opportunity to contribute to cutting-edge AI research and open-source initiatives.
- Combination of startup vitality and job stability.
- Simple, non-corporate work culture that respects individual beliefs.
- Continuous learning, growth, and support in an inclusive work environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
IT Engineer
150 000 - 250 000$
1 день назад
Systems Operations Support Engineer (AI Infrastructure)
90 000 - 160 000$
1 день назад
Technical Support Engineer (AI Infrastructure)
90 000 - 150 000$
1 день назад
Network Solutions Engineer (Robotics)
1 день назад
Infrastructure Engineer (AI Hardware)
150 000 - 250 000$
xAI
6 дней назад