Skip to content
aitrainer.work - AI Training Jobs Platform
STEM terac

Software Engineers: Paid Code Review for AI Agent Evaluation

Terac • United States, Canada

Education

Not stated

Type

Hourly

Pay Rate

$65/hr

Listed

Today

Apply Now →

About this role

From the Terac listing

What We're Researching

We're running a paid study on the coding environments and programming tasks used to benchmark artificial intelligence agents. Creating robust evaluation harnesses ensures that AI models are tested against realistic software engineering scenarios. This work directly feeds into improving how autonomous agents handle complex coding objectives.

How It Works

During this remote session, you will review a series of coding tasks and their corresponding evaluation harnesses. You will assess whether the programming challenges accurately reflect real-world software engineering problems. We will ask you to verify the logic, test cases, and overall structure of the environments provided. Your technical feedback will be captured through a guided conversation and screen-sharing exercises.

Who This Is For

We are looking for practicing software engineers with strong backgrounds in building and testing complex systems. Candidates should have direct experience writing test suites, evaluation harnesses, or comprehensive code reviews. We welcome backend engineers, full-stack developers, machine learning engineers, and software architects.

Why this role

This Software Engineers: Paid Code Review for AI Agent Evaluation position pays $65/hr. The bar for it is solid STEM knowledge, and the AI training workflow itself gets taught on the job.

Talent pool

We're light on STEM candidates

We've matched 65 people with a STEM background against 726 STEM listings we've tracked, so most go out without one. Set up a profile and we'll consider you for a role like this one.

Set up your profile

Skills and categories

Explore other opportunities in related specializations:

Related jobs

Terac

Browse All Jobs from Terac

Discover more opportunities on Terac that match your skills and interests.

View All Terac Jobs →

Verified Reviews

Loading reviews…

Community Reviews

Loading reviews…
💬

Share your experience with Terac

Help other candidates make better decisions by leaving a review.

Sign in to leave a review

Common questions

Is Terac legitimate?

Yes. Terac is a funded, US-based marketplace with real AI-lab and market-research clients, not a task-mill. It verifies your identity and professional background before matching you to any paid work, which is stricter screening than most open task queues use.

How does Terac decide who gets matched to this listing?

You complete a short AI-driven interview, identity verification, and a domain-specific screening test once. After that, Terac matches verified candidates to paid studies that fit their profile instead of running an open queue, and pay is released once your submission on this listing is checked against the requirements above.

Is this kind of AI-training work on Terac steady, or does it come and go?

Project-based, not steady. Terac opens listings like this in bursts when an AI lab requests a specific batch, categorizing data, writing or validating coding tasks, creating expert-level problems, and closes them once the batch is filled or complete. Treat it as recurring gig income you requalify for each time, not a standing job.

Why does pay vary so much across Terac's AI-training listings?

Pay tracks verified expertise, not a platform-wide rate. Generalist labeling sits at a few dollars an hour; work gated behind a passed domain screening, like software engineering or STEM problem creation, pays task bounties well over $100. Check this listing's own rate above rather than assuming other Terac listings pay the same.

Do I need a PhD for academic-niche AI training roles?

For the top pay tiers, a PhD or current enrollment is usually expected. But the domain assessment is what decides it: if you can solve the problems, the degree becomes secondary.

Is academic-niche AI training a continuous job?

Usually not. Work runs in campaigns, like training a model specifically on quantum mechanics, that might last a few weeks. Treat it as a high-paying fellowship or grant rather than a permanent daily job.

What does STEM work look like for a Software Engineers: Paid Code Review for AI Agent Evaluation?

Tasks here are scoped to STEM, not generic labeling. As a Software Engineers: Paid Code Review for AI Agent Evaluation, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to STEM) rather than following a one-size-fits-all rubric. If you don't have hands-on STEM background, this is likely not the right listing to start with.

What specific skills does this listing call for?

Coding and Expert are named directly in the listing. If you don't have hands-on experience with these, expect the screening process to test for them directly rather than accepting adjacent experience as a substitute.

How much does this specific role pay?

This listing is posted at $65/hr, an hourly rate. Pay can change between when we last checked the listing and when you apply, so confirm the current number on the platform's own application page before committing time.

What happens when I click Apply on this listing?

You'll be taken to Terac's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.

Can I apply from outside the United States or Canada?

This specific role is open only to people based in the United States and Canada. If you are somewhere else, applying is unlikely to lead to an offer even if you pass the assessment, because the restriction is usually about where the work can legally be contracted rather than your skills. Read the full description for any tax-residency or right-to-work caveats before you apply, since they can differ by country.