Назад
Company hidden
3 часа назад

AI Resident - Learning From Videos (LFV)

45 - 60$
Формат работы
hybrid
Тип работы
fulltime
Грейд
trainee
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Resident - Learning From Videos (LFV) (Multimodal AI and Robotics): Developing foundation models and algorithms for video, multimodal, and 4D reasoning in physical and embodied applications with an accent on computer vision, large-scale training, and spatio-temporal understanding. Focus on building world models, processing large-scale multimodal datasets, and solving zero-shot generalization challenges for real-world robotics tasks.

Location: Hybrid in Los Altos, California, United States

Salary: $45–$60 per hour for California-based roles.

Company

hirify.global develops AI, robotics, driving, and materials-science technologies focused on improving human mobility, independence, and quality of life.

What you will do

  • Develop, integrate, and deploy multimodal and 4D reasoning algorithms for physical applications.
  • Process large-scale training datasets, including streaming, online, and continual-learning data.
  • Develop machine learning and computer vision solutions for embodied robotics and real-world task performance.
  • Collaborate with robotics and machine learning researchers and engineers on theoretical and practical research needs.
  • Produce maintainable code for internal use and open-source scientific projects.
  • Contribute to research publications and technical reports.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, Robotics, or a related technical field.
  • Strong background in computer vision applied to robotics and embodied systems.
  • Research experience demonstrated through publications, technical projects, open-source contributions, or equivalent industry research experience.
  • Strong communication skills, collaborative mindset, and ability to learn quickly.
  • Interest in physical assistance innovation, dexterous manipulation, and human-robot collaboration for older adults and people in need.

Nice to have

  • Experience with spatio-temporal 4D computer vision, multi-view geometry, 3D/4D reconstruction, video generation, self-supervised learning, or occlusion reasoning.
  • Experience training large-scale multimodal deep learning models, extending context length, optimizing attention, or using distributed computing.
  • Experience applying machine learning and computer vision to embodied applications.

Culture & Benefits

  • Research-focused, year-long AI Residency integrated into the Learning From Videos robotics team.
  • Collaboration with Toyota affiliates and opportunities to contribute to research publications and technology productization.
  • Medical, dental, and vision insurance.
  • Paid time off, including holiday pay and sick time.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →