Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Field Data Center Hardware Engineer (AI): Deploying and bringing NVIDIA H200 and B200 GPU platforms into production across high-density data center environments with an accent on hardware validation, Linux systems, networking, and operational execution. Focus on troubleshooting hardware, firmware, OS, and network issues, coordinating repairs, driving root-cause analysis, and leading concurrent infrastructure rollouts.
Location: United States; on-site data center deployment and repair coordination required. Applicants must be authorized to work in the country in which they apply.
Salary: $112,000–$140,000 USD base annually, plus quarterly performance bonuses.
Company
Nebius builds a full-stack AI cloud platform for data and model training, inference optimization, and production deployment at scale.
What you will do
- Lead end-to-end deployment of GB-series racks in data center environments.
- Install, bring up, validate, and prepare NVIDIA H200 and B200 servers for production.
- Troubleshoot hardware, firmware, Linux OS, and networking issues across GPU-dense systems.
- Develop Linux-based hardware health-check and diagnostic scripts and maintain deployment runbooks.
- Coordinate on-site repairs, part replacements, vendor escalations, and root-cause analysis.
- Lead deployment timelines and provide technical guidance to on-site engineers and technicians.
Requirements
- Hands-on experience deploying and operating data center infrastructure.
- Strong knowledge of GPU-dense systems, high-density rack deployments, and NVIDIA H-series platforms.
- Solid Linux troubleshooting and scripting experience.
- Ability to diagnose issues across hardware, operating system, firmware, and network layers.
- Experience coordinating field repairs, working with hardware vendors, and leading technical teams or field operations.
- Authorization to work in the United States is required.
Nice to have
- Experience deploying AI or HPC clusters at scale.
- Familiarity with automated provisioning and infrastructure lifecycle systems.
- Experience in hardware qualification, burn-in testing, or factory validation.
- Experience supporting rapid infrastructure expansion.
- Exposure to ARM-based or heterogeneous compute environments.
Culture & Benefits
- Collaboration with globally distributed engineering and operations teams.
- Company-paid medical, dental, and vision coverage for employees and families.
- 401(k) plan with up to 4% company match and immediate vesting.
- Paid parental leave and company-paid disability and life insurance.
- Career growth, learning opportunities, flexibility, ownership, and work on AI infrastructure projects.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
AI Accelerator Compute Architect (AI)
120 000 - 275 000$
1 день назад
Hardware Design Engineer (AI Infrastructure)
180 000 - 280 000$
SandboxAQ
4 дня назад
Staff Forward Deployed Engineer (AI)
118 400 - 222 000$
1 день назад
Staff Engineer, Wafer Quality Assurance (AI)
102 500 - 136 600$
7 дней назад
Senior Supplier Quality Engineer (AI Infrastructure)
95 000 - 110 000$
Anthropic
4 дня назад
Data Center Mechanical Engineer
320 000 - 405 000$