×
Register Here to Apply for Jobs or Post Jobs. X

LATAM Software Engineers: Coding Tasks AI Evaluation

Job in Washington, District of Columbia, 20022, USA
Listing for: Remote Jobs
Full Time position
Listed on 2026-10-03
Job specializations:
  • Software Development
    Software Testing, AI Engineer (Applied/Software), AI QA / Validation Engineer, Software Engineer
Salary/Wage Range or Industry Benchmark: 40 USD Hourly USD 40.00 HOUR
Job Description & How to Apply Below
Position: LATAM Software Engineers: Coding Tasks for AI Evaluation

WHAT WE'RE RESEARCHING

We're running a paid study on the effectiveness of coding tasks used to evaluate AI agents. Our team is building a comprehensive suite of programming environments designed to test complex software capabilities. We want to ensure these evaluation frameworks are realistic, accurate, and properly calibrated.

How It Works

You will review a series of proposed coding tasks and assess their suitability for testing AI agents. We will ask you to verify the logic of the evaluation harnesses and provide feedback on their realistic application. You will share your screen to walk through the environments and point out potential flaws or improvements. The session involves a mix of code review, technical discussion, and direct feedback on the task structures.

Who

This Is for

We are looking for software engineers based in Brazil and Argentina with strong technical backgrounds. You should have direct experience building, reviewing, or testing evaluation harnesses and programming tasks. We welcome backend developers, full-stack engineers, and quality assurance automation specialists.

WHAT YOU'LL DO
  • Review realistic programming tasks designed for AI evaluation
  • Assess the accuracy and structure of various evaluation harnesses
  • Walk us through your thought process while checking code validity
  • Provide actionable feedback on how to improve task complexity and realism
Who Should Apply
  • Professional experience as a software engineer or developer
  • Located in Brazil or Argentina
  • Familiarity with building or verifying coding tasks and test environments
  • Comfortable discussing technical architectures and evaluation frameworks
Compensation

$40 per hour

About Terac

Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Learn more at  or on You Tube at @jointerac

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary