5 дней назад
Reinforcement Learning Researcher (AI)
250 000 - 350 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Reinforcement Learning Researcher (AI): Developing frontier reinforcement learning models for scientific reasoning and tasks across materials and energy research with an accent on scalable RL algorithms, adaptive sampling, and training stability. Focus on designing controlled experiments, improving compute efficiency, and transferring small-scale findings to large-scale training runs.
Location: Menlo Park, California, United States; on-site
Base compensation: $250,000–$350,000 per year, plus equity
Company
Periodic Labs is an AI and physical sciences company developing models to accelerate breakthroughs in materials, energy, and related scientific fields.
What you will do
- Design experiments to study how reinforcement learning scales with compute, model size, data, and reward quality.
- Develop RL algorithms for policy optimization, advantage estimation, exploration, and credit assignment in long-horizon tasks.
- Build adaptive sampling and curriculum methods that adjust task difficulty, problem selection, and rollout counts.
- Investigate bias and stability in RL training, including importance-sampling corrections, policy staleness, and training–inference mismatch.
- Improve training and inference efficiency through experiments with hyperparameters, rollout counts, batch sizes, and update schedules.
Requirements
- Hands-on experience training LLMs with reinforcement learning.
- Ability to design small-scale RL setups that transfer to large-scale training runs.
- Experience implementing, debugging, and testing research ideas across a complex training stack.
- Strong attention to detail and a rigorous scientific approach to answering research questions.
- At least 5 years of experience.
- Bachelor’s degree or equivalent experience.
Culture & Benefits
- Work on frontier models for scientific reasoning and discovery.
- Collaborate in a rapidly growing, research-focused environment.
- Receive equity in addition to base compensation.
- Visa sponsorship is available, with assistance throughout the process.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Post-Training Engineer (AI)
300 000 - 350 000$
Microsoft AI
10 дней назад
Software Engineer (Reinforcement Learning)
142 800 - 274 800$
Baseten
9 дней назад
Post-Training Research Scientist (AI)
210 000 - 285 000$
Anthropic
8 дней назад
Research Engineer Scientist (AI)
350 000 - 500 000$
Resolution
12 дней назад
Research Scientist (AI Alignment)
230 000 - 930 000$
5 дней назад