12 дней назад
Software Engineers (AI Evaluation)
65$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Software Engineers (AI Evaluation) (Python/open-source): Reviewing and evaluating coding tasks derived from real GitHub issues and pull requests with an accent on reproducibility, testing, dependency management, and multi-file code changes. Focus on identifying flaky tests, ambiguous requirements, environment issues, and task-complexity problems while providing clear recommendations for acceptance, improvement, or exclusion.
Location: Remote within Argentina, Brazil, Chile, Colombia, Ecuador, Mexico, Portugal, Spain, or Uruguay
Salary: $65 per hour
Company
is recruiting software engineers for a specialized project evaluating real-world software engineering tasks from open-source repositories.
What you will do
- Review coding tasks derived from real GitHub issues, pull requests, bug fixes, and feature implementations.
- Assess whether problem statements, success criteria, tests, dependencies, and repository setup are clear and complete.
- Evaluate unit tests for correctness, coverage, robustness, and appropriate implementation flexibility.
- Identify flaky tests, missing dependencies, version conflicts, environment-specific behavior, and reproducibility issues.
- Assess task difficulty, multi-file and cross-module complexity, and real-world relevance.
- Provide written recommendations on whether tasks should be accepted, improved, or excluded.
Requirements
- 3+ years of professional software engineering experience.
- Strong experience with large, multi-file codebases, production-code maintenance, debugging, and pull request review.
- Strong understanding of unit testing, test coverage, Git, and GitHub-based development workflows.
- Experience with dependency management, environment setup, reproducibility, and evaluating technical problems.
- Ability to provide clear written technical feedback.
Nice to have
- Contributions to or maintenance of open-source projects.
- Experience with SWE-Bench, SWE-Bench Verified, or similar coding benchmarks.
- Experience with major Python open-source projects such as Django, Flask, scikit-learn, SymPy, matplotlib, requests, or pytest.
- Experience with Docker, CI/CD, pip, conda, dependency pinning, test fixtures, test isolation, or property-based testing.
- Experience with technical assessments, AI/ML evaluation, data curation, RLHF, or benchmark development.
Culture & Benefits
- Remote, part-time, project-based consulting engagement.
- Focus on software engineering, open-source code review, testing, and technical evaluation.
- Work on real-world repositories and professional software engineering problems.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 дней назад
AI Trainer (Computer Science)
8 - 65$
12 дней назад
Spanish (Rioplatense) Language Specialist (AI Trainer)
8 - 65$
12 дней назад
Spanish (Chile) Language Specialist - Freelance AI Trainer Project (AI)
8 - 65$
12 дней назад
Spanish (Mexico) Language Specialist (AI)
8 - 65$
13 дней назад
Spanish Language Specialist (Latin America) (AI Trainer)
6 - 65$
12 дней назад
Spanish (Andean — Peru/Bolivia/Ecuador) Language Specialist - AI Trainer
8 - 65$