Skip to content
aitrainer.work - AI Training Jobs Platform

Older listing: position may have been filled

This listing is no longer actively promoted, but you're still welcome to apply. Platforms often reopen roles or keep applications on file.

Data Science alignerr

Machine Learning Evaluation Specialist

Alignerr • Remote

Education

Not stated

Type

Hourly

Pay Rate

$200–$400/hr

Listed

181d ago

Apply opens Alignerr in a new tab.

Check Listing →

About this role

From the Alignerr listing

What You'll Do

  • Design complex, original machine learning problems rooted in your area of domain expertise
  • Create evaluation tasks that demand advanced knowledge well beyond standard ML pipelines
  • Draw from your own research experience to craft challenges that genuinely test highly capable AI models
  • Write clear problem statements, define evaluation criteria, and establish gold-standard solutions
  • Assess AI-generated solutions for correctness, creativity, and methodological rigor
  • Document problem difficulty, required domain knowledge, and expected failure modes
  • Collaborate asynchronously with a global team of researchers and engineers

About the Role

What if your years of hard-earned research expertise could directly shape the future of AI? We're looking for domain experts with deep machine learning knowledge to design evaluation challenges that push state-of-the-art AI systems to their limits — the kind of problems only a true specialist could craft. Your work won't sit in a drawer. It directly influences how the next generation of AI models are measured, trained, and improved.

  • Organization: Alignerr
  • Type: Hourly Contract
  • Location: Remote
  • Commitment: 10–40 hours/week

Who You Are

  • Graduate-level expertise (MS or PhD preferred) in a scientific or technical discipline that intersects with machine learning
  • Strong working knowledge of ML methods — model selection, feature engineering, evaluation metrics, and pipeline design
  • Deep familiarity with active, open research problems in your field
  • A sharp eye for where general ML knowledge breaks down and specialized domain insight becomes essential
  • Experience publishing or conducting original research is highly valued
  • Excellent written communication — you can articulate complex, nuanced problems with precision and clarity
  • Self-motivated and energized by intellectually demanding, independent work

Example Domains

We welcome experts from a wide range of fields, including but not limited to: If your domain sits at the frontier of ML research, we want to hear from you.

  • Computational biology, genomics, or bioinformatics
  • Climate science and environmental modeling
  • Medical imaging and healthcare ML
  • Materials science and computational chemistry
  • Astrophysics and signal processing
  • Natural language processing for low-resource or specialized corpora
  • Robotics, control theory, or reinforcement learning in complex environments
  • Financial modeling and quantitative analysis

Why Join Us

  • Work at the cutting edge — your challenges help define the boundaries of what AI can and cannot do
  • Make a real impact — your expertise directly shapes AI safety and evaluation research
  • Full autonomy — work on your own schedule, from anywhere in the world
  • Flexible commitment — scale hours up or down based on your availability
  • Ongoing opportunity — strong contributors are considered for contract extensions and deeper research involvement
  • Build your profile — establish yourself as a contributor to frontier AI development alongside top research labs

Why this role

What if your years of hard-earned research expertise could directly shape the future of AI? We're looking for domain experts with deep machine learning knowledge to design evaluation challenges that push state-of-the-art AI systems to their limits — the kind of problems only a true specialist could craft. Your work won't sit in a drawer. It directl

Skills and categories

Explore other opportunities in related specializations:

Related jobs

Alignerr

Browse All Jobs from Alignerr

Discover more opportunities on Alignerr that match your skills and interests.

View All Alignerr Jobs →

Verified Reviews

Loading reviews…

Community Reviews

Loading reviews…
💬

Share your experience with Alignerr

Help other candidates make better decisions by leaving a review.

Sign in to leave a review

Common questions

Does Alignerr have a trainer community?

Yes, and it's a genuine strength. Once you're assigned to a project, you join Slack channels where you can get rubric clarifications from admins and talk to other trainers. That kind of support is rare in AI training and matters most when guidelines are ambiguous or shift mid-project.

How hard is the Alignerr assessment?

Hard, and unforgiving. Alignerr uses TestGorilla for timed, role-specific tests: a blank coding environment for engineers, strict grammar and fact-checking for writers. Treat it as one shot. Failing or abandoning it typically locks you out of that role permanently, with no retake.

Is AI training work the same as traditional consulting?

No. Instead of client deliverables, you're given complex scenarios to evaluate: grading the AI's logic, correcting its hallucinations, and supplying expert-level reasoning it doesn't have on its own. The job is closer to teaching than consulting.

Why do these AI training roles pay so much?

Because general knowledge isn't what's being tested. The model already knows the basics; what it needs is expertise on edge cases, the rare, difficult, highly technical judgment calls only a senior professional in the field would make correctly.

What does Data Science work look like for a Machine Learning Evaluation Specialist?

Tasks here are scoped to Data Science, not generic labeling. As a Machine Learning Evaluation Specialist, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Data Science) rather than following a one-size-fits-all rubric. If you don't have hands-on Data Science background, this is likely not the right listing to start with.

What specific skills does this listing call for?

Expert is named directly in the listing. If you don't have hands-on experience with this, expect the screening process to test for it directly rather than accepting adjacent experience as a substitute.

How much does this specific role pay?

This listing is posted at $200–$400/hr, an hourly rate. The range reflects experience level and negotiated terms, not a placeholder, so where you land in it depends on your background and the assessment. Pay can change between when we last checked the listing and when you apply, so confirm the current number on the platform's own application page before committing time.

What happens when I click Apply on this listing?

You'll be taken to Alignerr's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.

What is the barrier to entry for Alignerr?

A difficult, timed technical assessment in your specific domain, like Python, physics, or language. Passing it is required before you're eligible for any paid projects.