5 дней назад
Senior AI Coding Agent Evaluator
100 - 200$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior AI Coding Agent Evaluator (AI/Coding Agents): Evaluating end-to-end interactions with modern coding agents such as OpenAI Codex, Claude Code, and Cursor with an accent on engineering judgment, response quality, and useful explanations. Focus on distinguishing strong from misleading model behavior, defining quality standards, and providing rigorous written and video feedback without reviewing every line of code.
Location: Remote, worldwide
Salary: $100–$200/hour, depending on experience and location
Company
connects experienced engineers with projects involving software development and evaluation of modern coding tools.
What you will do
- Evaluate AI-generated coding interactions from start to finish.
- Judge whether responses are useful, correct at a high level, and aligned with strong engineering judgment.
- Assess the quality of explanations, reasoning, preambles, and guidance—not only code syntax.
- Distinguish different levels of response quality and define what excellent interaction looks like.
- Provide clear, opinionated written feedback on what worked, failed, or felt misleading.
- Record short video explanations of evaluation decisions.
Requirements
- Senior, Staff, or Principal-level engineering experience, or equivalent.
- Strong background in TypeScript/JavaScript or Python.
- Hands-on experience with at least one of OpenAI Codex, Claude Code, or Cursor.
- Deep familiarity with modern AI-assisted development workflows.
- Strong written and spoken English at B2 level or above is required.
- Experience with prompt design or evaluation workflows, mentoring senior engineers, or defining engineering standards.
Nice to have
- Experience evaluating code without executing it or reviewing every line.
- Prior experience recording technical video explanations.
Culture & Benefits
- Contract engagement with flexible hours of 10–20 hours per week.
- Projects typically last from approximately two weeks to a few months.
- Ongoing opportunities are offered to evaluators who perform well.
- Start as soon as the take-home evaluation is completed and a project seat is available.
Hiring process
- Complete one take-home evaluation exercise.
- Submit a recorded Loom walkthrough.
- No interview is required.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
micro1
8 дней назад
Senior Software Engineer (AI)
50 - 100$
9 дней назад
Software Engineer (AI)
40 - 150$
8 дней назад
Latvian Language Specialist - Freelance AI Trainer Project (AI)
6 - 65$
8 дней назад
Malay Language Specialist (AI Trainer)
6 - 65$
8 дней назад
Korean Language Specialist (AI)
6 - 65$
8 дней назад
Malayalam Language Specialist (AI)
8 - 65$