LATAM Software Engineers: Coding Tasks for AI Evaluation
Terac • Argentina, Brazil
Education
Not stated
Type
Hourly
Pay Rate
$40/hr
Listed
1d ago
About this role
From the Terac listing
What We're Researching
We're running a paid study on the effectiveness of coding tasks used to evaluate AI agents. Our team is building a comprehensive suite of programming environments designed to test complex software capabilities. We want to ensure these evaluation frameworks are realistic, accurate, and properly calibrated.
How It Works
You will review a series of proposed coding tasks and assess their suitability for testing AI agents. We will ask you to verify the logic of the evaluation harnesses and provide feedback on their realistic application. You will share your screen to walk through the environments and point out potential flaws or improvements. The session involves a mix of code review, technical discussion, and direct feedback on the task structures.
Who This Is For
We are looking for software engineers based in Brazil and Argentina with strong technical backgrounds. You should have direct experience building, reviewing, or testing evaluation harnesses and programming tasks. We welcome backend developers, full-stack engineers, and quality assurance automation specialists.
Why this role
At $40/hr, this LATAM Software Engineers: Coding Tasks for AI Evaluation position pays for what you already know about STEM. The AI training side of the job is covered during onboarding.
Talent pool
We're light on STEM candidates
We've matched 65 people with a STEM background against 726 STEM listings we've tracked, so most go out without one. Set up a profile and we'll consider you for a role like this one.
Set up your profileSkills and categories
Explore other opportunities in related specializations:
Related jobs
Browse All Jobs from Terac
Discover more opportunities on Terac that match your skills and interests.
View All Terac Jobs →Verified Reviews
Community Reviews
Share your experience with Terac
Help other candidates make better decisions by leaving a review.
Sign in to leave a reviewLeave your review
Common questions
Is Terac legitimate?
Yes. Terac is a funded, US-based marketplace with real AI-lab and market-research clients, not a task-mill. It verifies your identity and professional background before matching you to any paid work, which is stricter screening than most open task queues use.
How does Terac decide who gets matched to this listing?
You complete a short AI-driven interview, identity verification, and a domain-specific screening test once. After that, Terac matches verified candidates to paid studies that fit their profile instead of running an open queue, and pay is released once your submission on this listing is checked against the requirements above.
Is this kind of AI-training work on Terac steady, or does it come and go?
Project-based, not steady. Terac opens listings like this in bursts when an AI lab requests a specific batch, categorizing data, writing or validating coding tasks, creating expert-level problems, and closes them once the batch is filled or complete. Treat it as recurring gig income you requalify for each time, not a standing job.
Why does pay vary so much across Terac's AI-training listings?
Pay tracks verified expertise, not a platform-wide rate. Generalist labeling sits at a few dollars an hour; work gated behind a passed domain screening, like software engineering or STEM problem creation, pays task bounties well over $100. Check this listing's own rate above rather than assuming other Terac listings pay the same.
What does asynchronous AI training work mean in practice?
No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.
What does STEM work look like for a LATAM Software Engineers: Coding Tasks for AI Evaluation?
Tasks here are scoped to STEM, not generic labeling. As a LATAM Software Engineers: Coding Tasks for AI Evaluation, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to STEM) rather than following a one-size-fits-all rubric. If you don't have hands-on STEM background, this is likely not the right listing to start with.
What specific skills does this listing call for?
Coding and Expert are named directly in the listing. If you don't have hands-on experience with these, expect the screening process to test for them directly rather than accepting adjacent experience as a substitute.
How much does this specific role pay?
This listing is posted at $40/hr, an hourly rate. Pay can change between when we last checked the listing and when you apply, so confirm the current number on the platform's own application page before committing time.
What happens when I click Apply on this listing?
You'll be taken to Terac's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.
Can I apply from outside Argentina or Brazil?
This specific role is open only to people based in Argentina and Brazil. If you are somewhere else, applying is unlikely to lead to an offer even if you pass the assessment, because the restriction is usually about where the work can legally be contracted rather than your skills. Read the full description for any tax-residency or right-to-work caveats before you apply, since they can differ by country.