Назад
Company hidden
обновлено 1 день назад

Principal AI Research Scientist Post-Training (AI)

Формат работы
remote (Global)/onsite
Тип работы
fulltime
Грейд
senior
Английский
c1
Страна
UK/US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Principal AI Research Scientist Post-Training (AI): Developing and optimizing post-training methods for foundation models with an accent on RLHF, preference optimization, and long-horizon reasoning. Focus on grounding reinforcement learning in physics and engineering constraints to improve model reliability and controllability.

Location: Remote (US, Canada, EU) or onsite in San Francisco, London, or Toronto

Company

hirify.global provides software tools for architecture, engineering, construction, manufacturing, and media industries, leveraging high-fidelity physics simulation and CAD kernels for AI research.

What you will do

  • Lead post-training model development including RLHF, preference optimization, and agentic systems.
  • Develop novel algorithms to enhance model reliability, controllability, and alignment.
  • Design and execute experiments to shape model behavior and reasoning quality.
  • Build scalable and reproducible post-training workflows in collaboration with infrastructure teams.
  • Create evaluation frameworks for long-horizon reasoning, tool use, and real-world workflow completion.
  • Contribute to high-impact publications, patents, and external research visibility.

Requirements

  • PhD or equivalent industry research experience in ML, RL, AI, or related fields.
  • Deep expertise in reinforcement learning for foundation models (RLHF, RLAIF, DPO, PPO).
  • Proven track record of leading or mentoring technical research teams.
  • Strong publication record at leading ML/AI venues (e.g., NeurIPS, ICML, ICLR).
  • Experience at a frontier model lab or advanced applied AI organization.
  • Must be based in the US, Canada, or EU for remote work.

Culture & Benefits

  • Opportunity to work at the intersection of AI research and real-world professional workflows.
  • Direct pipeline from research advances to large-scale product impact.
  • Collaborative environment with leading academic and industry labs.
  • Inclusive culture focused on diversity and equal opportunity.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →