Назад
Company hidden
2 часа назад

Infrastructure Engineer (AI Inference)

200 000 - 400 000$
Формат работы
remote (только USA)/onsite
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Infrastructure Engineer (AI Inference): Building distributed systems and foundational infrastructure that enable vLLM to serve AI models across thousands of accelerators with an accent on performance, low latency, and reliability. Focus on designing high-performance systems, debugging complex distributed issues, and integrating GPU interconnects and disaggregated inference architecture.

Location: San Francisco, California; remote within the United States may be considered for exceptional candidates

Salary: $200,000–$400,000 per year plus equity

Company

hirify.global develops infrastructure around vLLM to make AI inference cheaper, faster, and more reliable.

What you will do

  • Build distributed systems and foundational infrastructure for serving models through vLLM at global scale.
  • Design systems that support inference across thousands of accelerators with low latency and high reliability.
  • Develop high-performance systems using Rust, Go, or C++.
  • Work with network protocols, high-performance I/O, GPU programming models, and memory hierarchies.
  • Debug complex distributed systems issues and improve system performance and reliability.

Requirements

  • Bachelor’s degree or equivalent experience in computer science, engineering, or a similar field.
  • Strong systems programming skills in Rust, Go, or C++.
  • Experience designing and building high-performance distributed systems at scale.
  • Understanding of network protocols and high-performance I/O.
  • Ability to debug complex distributed systems issues.
  • Work from the United States is required for remote candidates; the primary location is San Francisco.

Nice to have

  • Experience with ML serving infrastructure and disaggregated inference architecture.
  • Familiarity with GPU programming models, memory hierarchies, and GPU interconnects such as NVLink, InfiniBand, or RoCE.
  • Experience supporting large-scale model training or inference environments.
  • Track record of improving system reliability and performance at scale.

Culture & Benefits

  • Health, dental, and vision benefits.
  • 401(k) company match.
  • Equity included in the compensation package.
  • Visa sponsorship is available on a case-by-case basis.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →