SWE-Bench AI Task Auditor - Freelance AI Trainer Project
Project Overview We are sourcing experienced technical specialists in Software Engineering (SWE-Bench) to audit tasks used to train and evaluate AI systems. The objective of this project is to ensure that AI training workflows and tasks are technically rigorous, practical, and highly accurate within your specific area of expertise.
Project Deliverables & Scope
-
Task Evaluation: Assess whether software engineering tasks are technically accurate, realistic, solvable, reproducible, and supported by reliable tests and evaluation criteria.
-
Technical Auditing & Feedback: Provide clear, actionable feedback on any identified codebase integration issues, test failures, or logic errors within the tasks.
Required Expertise
-
Demonstrable professional experience and deep technical knowledge in software engineering (including complex codebase navigation, real-world application development, and SWE-Bench standards). (Note: Candidates will be considered for one specialty based on their experience; expertise across other domains is not required).
-
Strong analytical and problem-solving skills to rigorously test, troubleshoot, and evaluate complex technical scenarios.
We offer a pay range of $70 - $100 per hour, with the exact rate determined after evaluating your experience, expertise, and geographic location. Final offer amounts may vary from the pay range listed above. As a contractor you’ll supply a secure computer and high‑speed internet; company‑sponsored benefits such as health insurance and PTO do not apply.
Employment type: Freelance Contract
Workplace type: Remote
Apply for this job
*
indicates a required field

