Назад
Company hidden
4 часа назад

Principal System Software Engineer, AI Inference Execution

195 000 - 285 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal System Software Engineer, AI Inference Execution (AI/C++/Python): Building and scaling deployment infrastructure and next-generation software for an AI compute engine with an accent on system software, distributed high-performance design, and hardware-software co-design. Focus on optimizing inference execution, integrating ML and compiler software with hardware, and deploying large-scale AI workloads.

Location: Hybrid, onsite at the Santa Clara, CA headquarters 3 days per week

Salary: $195,000–$285,000 annual base compensation, plus equity and bonus opportunities

Company

hirify.global develops software and hardware innovations for generative AI and AI computing.

What you will do

  • Productize the software stack for the AI compute engine.
  • Develop, enhance, and maintain next-generation AI deployment software.
  • Build and scale deployment infrastructure within tight development timelines.
  • Optimize hardware-software co-design trade-offs across the full-stack toolchain.
  • Collaborate with system software, machine learning, compiler, and hardware experts.

Requirements

  • 12+ years of industry software development experience with a bachelor's degree, or 6+ years with a master's degree in a relevant field.
  • Strong knowledge of system software, data structures, computer architecture, and machine learning fundamentals.
  • Proficiency in C, C++, and Python in a Linux environment.
  • Experience designing and implementing distributed, high-performance software.
  • Strong ownership, leadership, collaboration, and self-motivation.

Nice to have

  • Experience with inference servers and model-serving frameworks such as TensorRT-LLM, vLLM, or SGLang.
  • Experience with PyTorch, TensorFlow, ONNX Runtime, or TensorRT.
  • Experience with NCCL, OpenMPI, Kubernetes, Ray, or other MLOps tools.
  • Experience deploying LLM, VLM, or NLP workloads on distributed systems.
  • Startup, small-team, cloud-provider, or AI-compute company experience.

Culture & Benefits

  • Collaborative and inclusive environment focused on respect, humility, and direct communication.
  • Medical, dental, and vision coverage.
  • 401(k) and an employee rewards package supporting overall wellbeing.
  • Equity and performance-based bonus opportunities.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →