Company logo hidden

SWE-Bench AI Task Auditor - Freelance AI Trainer Project

Unlock employer United Arab Emirates Direct to Company 1 hour ago · 30 Sep 2026

Financial

  • Estimate: $124k - $124k*
  • Zero income tax location

Accessibility

  • Fully Remote
  • Visa Provided

Requirements

  • Experience: Senior
  • English: Professional

Position

About the Role
This freelance opportunity is designed for experienced software engineers who want to contribute their technical expertise to the development and evaluation of AI systems. You will review software engineering tasks used to train and assess advanced AI models. Your work will focus on ensuring tasks are technically accurate, realistic, reproducible, and appropriately tested. You will investigate codebase integrations, test failures, logic issues, and other technical challenges. The role offers flexibility through a fully remote, project-based working model. You will apply practical software engineering judgment rather than simply following predefined checks. Your feedback will directly help improve the quality and reliability of AI training workflows.

Ready to apply for roles like this?

Unlock the company name and direct application link. Subscribers get instant access to fresh jobs across Dubai, Abu Dhabi and Riyadh, many with visa support.

Unlock employer & apply directly

Accountabilities

  • Evaluate software engineering tasks for technical accuracy, realism, solvability, reproducibility, and alignment with SWE-Bench standards.
  • Review task codebases, integrations, tests, and evaluation criteria to identify potential technical weaknesses.
  • Rigorously test and troubleshoot complex technical scenarios to determine whether tasks function as intended.
  • Investigate codebase integration problems, test failures, logic errors, and other implementation issues.
  • Provide clear, precise, and actionable feedback that enables task creators to correct identified problems.
  • Apply professional software engineering judgment to assess whether tasks reflect realistic development scenarios.
  • Help maintain a high standard of technical rigor and accuracy across AI training and evaluation workflows.

Requirements

  • Demonstrable professional experience in software engineering, including experience navigating complex codebases and developing real-world applications.
  • Strong knowledge of software engineering principles and familiarity with SWE-Bench-style tasks and evaluation workflows.
  • Strong analytical and problem-solving abilities, with the capacity to investigate complex technical scenarios systematically.
  • Experience troubleshooting code, diagnosing test failures, and identifying underlying logic or integration issues.
  • Ability to evaluate technical work objectively and communicate findings clearly and constructively.
  • Strong attention to detail and a rigorous approach to technical validation.
  • Ability to work independently and manage project-based assignments effectively.
  • Deep expertise in one relevant software engineering specialty is sufficient; expertise across multiple domains is not required.
  • A secure computer and reliable, high-speed internet connection suitable for remote technical work.

Benefits

  • Fully remote freelance contract.
  • Flexible project-based working environment.
  • Opportunity to contribute directly to the training and evaluation of AI systems.
  • Ability to apply real-world software engineering expertise to challenging technical tasks.
  • Compensation of $60/hour, with the final rate determined based on experience, expertise, and geographic location.
  • No company-sponsored health insurance, PTO, or other employee benefits, as this is a freelance contractor position.
Apply Direct

Jobs you might like   View all jobs

About Internet Marketplace Platforms Company

Company details are hidden. Subscribe to view full company profile.

Ready to apply for this role?

Apply Direct