Назад
Company hidden
2 месяца назад

AI Research Engineer

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
India
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Research Engineer (Autonomous Coding Agents): Advancing autonomous coding agents by defining rigorous quality metrics, building evaluation datasets, and running training and post-training experiments with an accent on agent behavior, reliability, and production-grounded measurement. Focus on designing evaluations for reasoning, planning, tool use, code correctness, security, and long-horizon execution, while turning research advances into shipped improvements.

Location: Bangalore, India

Company

hirify.global builds autonomous coding agents that generate, test, and deploy production applications from plain-language intent at global scale.

What you will do

  • Shape the core architecture and foundational design of autonomous coding agents.
  • Analyze agent behavior across real-world usage and convert success and failure patterns into rigorous quantitative measurements.
  • Design and ship evaluations covering reasoning, planning, tool use, code correctness, long-horizon execution, security, and reliability.
  • Build evaluation datasets, methodologies, validation processes, and dashboards for detecting regressions.
  • Run training and post-training experiments, including supervised fine-tuning, RLHF/RLAIF, DPO, distillation, reward modeling, prompt optimization, and judge-model calibration.
  • Own research work end-to-end, from hypothesis and experiment design through rollout and post-launch measurement.

Requirements

  • 5–8 years of AI experience, including model training and fine-tuning or rigorous evaluation and measurement systems.
  • Fluency in Python for training pipelines, evaluation harnesses, data processing, and statistical analysis; Go is a plus.
  • Experience with transformers, RLHF, DPO, reinforcement learning for agents, evaluation frameworks, prompt optimization, judge models, and agent frameworks.
  • Ability to reason about noise floors, confounds, distribution shift, judge bias, selection effects, and statistical validity.
  • Independence in scoping work, making decisions in subjective systems, communicating reasoning, and owning outcomes.
  • Ability to move quickly while maintaining rigorous and honest measurement.

Culture & Benefits

  • Lunch and dinner provided daily.
  • Family insurance with coverage of 3 lakhs for the employee and family.
  • Unlimited paid time off.
  • Flexible working hours.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →