4 часа назад
Member of Technical Staff (Research)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Member of Technical Staff (Research) (LLMs): Designing and building a post-training stack that enables continual learning from real production workflows, with an accent on telemetry, custom LLMs and agents, and evaluation. Focus on building RL environments, designing novel algorithms, and running end-to-end experiments across data, training, and model serving.
Location: San Francisco, United States; on-site
Company
is a research and product lab creating a platform for continual learning that enables companies to continuously post-train large-scale agentic models from product-use signals.
What you will do
- Design and build the post-training stack for continual learning from real production workflows.
- Shape product telemetry into learnable signals for model improvement.
- Train and serve custom large language models and agents.
- Design novel algorithms and own end-to-end experiments across data, training, and evaluation.
- Make foundational technical and product decisions with founders and design partners.
- Help define the architecture of the continually learning intelligence platform.
Requirements
- Experience training or serving large language models.
- Experience building RL environments or evaluation suites for LLMs.
- Proficiency in PyTorch, JAX, or similar machine learning frameworks.
- Strong experimental design skills.
- Master’s degree or research experience.
- Ability to work on-site in San Francisco.
Nice to have
- Background in pre-training or post-training reinforcement learning for LLMs.
- Experience with high-performance computing or large-scale clusters.
- Contributions to open-source machine learning research or infrastructure.
- Technical creativity demonstrated through research, open-source work, or side projects.
- PhD or final-year PhD work focused on AI or reinforcement learning open-source projects.
Culture & Benefits
- Work directly with founders and design partners.
- Collaborate with researchers from DeepMind, OpenAI, Meta, Apple, Amazon, and Scale.
- Contribute to the technical culture and architecture of a research-driven company.
- Build infrastructure for leading AI companies, including Decagon, Clay, Harvey, Mercor, and Rogo.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 часов назад
Research Scientist (Reinforcement Learning)
200 000 - 335 000$
12 часов назад
Member of Technical Staff, Post-Training (AI)
180 000 - 450 000$
4 часа назад
Founding Member of the Technical Staff (AI)
7 часов назад
Member of Technical Staff (AI)
3 часа назад
RL Environment Research Engineer (AI)
11 часов назад