Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Machine Learning Research Scientist, Evaluations (AI) (LLM evaluation and post-training): Building rigorous benchmarks and diagnostic methods for frontier language and multimodal models with an accent on failure-mode analysis, evaluation design, and post-training interventions. Focus on analyzing reasoning, robustness, and alignment failures, connecting them to training data and interventions, and publishing research findings.
Company
Scale AI develops high-quality data and full-stack technologies that support leading AI models and enterprise and government AI applications.
What you will do
- Analyze frontier LLM and agent behavior to identify, characterize, and diagnose capability gaps, reasoning errors, robustness issues, and alignment failures.
- Design and build benchmarks and evaluation methods for text and multimodal models.
- Apply SFT, RLHF, preference modeling, reward modeling, and instruction-tuning expertise to connect model failures with data and training interventions.
- Collaborate with researchers, engineers, and foundation model labs to establish evaluation-driven AI development practices.
- Translate failure analysis into technical and strategic input for future generative AI models.
- Publish research findings at major AI conferences and journals.
Requirements
- Ph.D. or Master's degree in Computer Science, Machine Learning, AI, or a related field.
- Deep understanding of deep learning, reinforcement learning, and large-scale model fine-tuning.
- Experience with LLM evaluation, benchmark development, and post-training techniques such as RLHF, preference modeling, or instruction tuning.
- Excellent written and verbal communication skills.
- Published machine learning research in major conferences or journals.
- Previous experience in a customer-facing role.
Culture & Benefits
- Base salary, equity, and benefits may be included for eligible roles.
- Comprehensive health, dental, and vision coverage.
- Retirement benefits, a learning and development stipend, and generous paid time off.
- Additional benefits may include a commuter stipend.
- Inclusive equal opportunity workplace with reasonable accommodations available during the application and recruiting process.
Hiring process
- Compensation is determined during the interview process based on work location, skills, experience, qualifications, interview performance, and education or training.
- Candidates must wait 90 days before being reconsidered for the same role.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
12 часов назад
Machine Learning Research Scientist, Evaluations (AI)
180 600 - 225 750$
4 часа назад
Machine Learning Research Scientist, Post-Training (AI)
180 600 - 225 750$
7 дней назад
Senior Applied Scientist (AI)
7 дней назад
Machine Learning Engineer, Research (Healthcare AI)
225 000 - 300 000$
3 дня назад
Senior Machine Learning Engineer, Public Sector (AI)
225 750 - 282 450$
6 дней назад