Назад
Company hidden
4 часа назад

Member of Technical Staff (Research)

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Member of Technical Staff (Research) (LLMs): Designing and building a post-training stack that enables continual learning from real production workflows, with an accent on telemetry, custom LLMs and agents, and evaluation. Focus on building RL environments, designing novel algorithms, and running end-to-end experiments across data, training, and model serving.

Location: San Francisco, United States; on-site

Company

hirify.global is a research and product lab creating a platform for continual learning that enables companies to continuously post-train large-scale agentic models from product-use signals.

What you will do

  • Design and build the post-training stack for continual learning from real production workflows.
  • Shape product telemetry into learnable signals for model improvement.
  • Train and serve custom large language models and agents.
  • Design novel algorithms and own end-to-end experiments across data, training, and evaluation.
  • Make foundational technical and product decisions with founders and design partners.
  • Help define the architecture of the continually learning intelligence platform.

Requirements

  • Experience training or serving large language models.
  • Experience building RL environments or evaluation suites for LLMs.
  • Proficiency in PyTorch, JAX, or similar machine learning frameworks.
  • Strong experimental design skills.
  • Master’s degree or research experience.
  • Ability to work on-site in San Francisco.

Nice to have

  • Background in pre-training or post-training reinforcement learning for LLMs.
  • Experience with high-performance computing or large-scale clusters.
  • Contributions to open-source machine learning research or infrastructure.
  • Technical creativity demonstrated through research, open-source work, or side projects.
  • PhD or final-year PhD work focused on AI or reinforcement learning open-source projects.

Culture & Benefits

  • Work directly with founders and design partners.
  • Collaborate with researchers from DeepMind, OpenAI, Meta, Apple, Amazon, and Scale.
  • Contribute to the technical culture and architecture of a research-driven company.
  • Build infrastructure for leading AI companies, including Decagon, Clay, Harvey, Mercor, and Rogo.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →