Назад
Company hidden
1 день назад

Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

170 000 - 260 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving) (AI/autonomous driving): Building and deploying Vision-Language-Action models that generate high-level driving decisions and trajectory guidance for autonomous trucks with an accent on large-scale training, on-vehicle validation, and onboard model compression. Focus on designing training and evaluation pipelines, applying SFT and RL post-training, and transitioning large reasoning models to onboard compute.

Location: Santa Clara, California, United States; hybrid workplace

Salary: $170,000–$260,000 per year

Company

hirify.global is a Physical AI company developing AI-based virtual driver software for factory-built autonomous trucks, with operations in the United States and Europe.

What you will do

  • Design, train, and evaluate Vision-Language-Action models for high-level driving decisions and trajectory guidance.
  • Own a VLA workstream across data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines with rigorous metrics for driving-context performance.
  • Develop distillation and compression methods for deploying reasoning models on onboard compute.
  • Apply supervised fine-tuning and reinforcement learning post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to move models from research into production.

Requirements

  • Master’s degree minimum in computer science, electrical engineering, mathematics, statistics, or a related field; Ph.D. preferred.
  • 3+ years of implementing and training models with PyTorch, TensorFlow, or JAX.
  • Hands-on experience training vision-language or vision-language-action models.
  • Experience with model training, evaluation, and production deployment.
  • Strong understanding of vision-language and VLA models, diffusion, flow matching, and transformers.
  • Experience with large-scale or distributed model training.

Nice to have

  • Experience with model distillation, quantization, inference optimization, ONNX/TensorRT, mixed precision, or custom kernels.
  • Experience with supervised fine-tuning and reinforcement learning post-training of large multimodal models.
  • Experience with camera, LiDAR, and radar data, autonomous driving, or ADAS.
  • Publications at leading computer vision, machine learning, or robotics venues.

Culture & Benefits

  • Work in an innovative, future-oriented autonomous driving environment.
  • Professional development opportunities.
  • Catered free lunch, unlimited snacks, and beverages.
  • Competitive salary and benefits package, including a 401(k) plan.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →