Pay Rates in the AI Evaluation Economy (2026): A Cross-Platform Benchmark of 2,628 Active Listings
We analyzed 2,628 active job listings across 8 AI evaluation platforms to benchmark hourly pay rates across platform tiers, role tracks, and pay distribution spreads.
Abstract. We conducted a quantitative pay-rate study across 2,628 active job listings on AI evaluation platforms, normalizing disclosed pay rates to USD per hour. Across 2,462 usable rate observations clearing baseline filters, median hourly rates split sharply into distinct platform tiers: high-end specialist platforms like Mercor ($80.00/hr median) and Alignerr ($80.00/hr median), mid-tier domain evaluation platforms like SME Careers ($55.00/hr median) and Micro1 ($48.50/hr median), and global developer platforms like Turing ($27.00/hr median). We report median, mean, interquartile range (P25βP75), and min/max pay benchmarks alongside explicit sample sizes ($n \ge 10$).
Key takeaways for job seekers & candidates
- Specialist Track Pay Premium: Mercor ($80.00/hr median) and Alignerr ($80.00/hr median) command top rates in the evaluation market, heavily driven by expert roles in specialized law, medicine, and software architecture.
- Standardized Global Rates: SME Careers ($55.00/hr median) and Micro1 ($48.50/hr median) concentrate in a predictable $30.00β$75.00/hr range for general domain training and bilingual evaluation.
- Entry-Level / Global Cost Tiers: Turing ($27.00/hr median) benchmarks at the lower end of technical AI evaluation, reflecting a global contractor pool optimized for international software development.
Methodology
Scope & Exclusions. The dataset includes active job listings across primary employer platforms as of July 2026 ($N = 2,628$ total listings). Non-employer re-post aggregators (Ethos) and task-based research sites (Terac) were excluded. Review platforms (Handshake, Mindrift, Outlier) were excluded from this active listings pool as they do not feed live listing data directly. Platforms lacking structured pay data (such as AfterQuery) were excluded from rate tables.
Rate Normalization to USD/Hour. Stated pay ranges were normalized to USD/hr. For listings with ranges, the midpoint was calculated; zero-floor placeholders were treated as missing data to avoid skewing range midpoints. Non-hourly rates were converted using standard benchmark commitments (30 hrs/week). Non-USD listings were excluded to eliminate foreign exchange volatility. Only platforms meeting an active sample threshold of $n \ge 10$ are reported as primary findings.
Note on Mercor's weight in the pool. Mercor accounts for nearly half of all usable observations ($n = 1{,}230$ of 2,462). Because Mercor mixes high-end domain-expert roles with a large volume of generalist task work, its scale means Mercor-specific pay dynamics have outsized influence on the overall market picture presented here β the tier breakdown below should be read platform-by-platform rather than as a single blended market rate.
Finding 1: Platform Pay Benchmarks
Pay rates across the AI training labor market exhibit significant structural variance depending on platform target candidates and sourcing models.
| Platform | Total Listings | Usable ($n$) | Median $/hr | Mean $/hr | P25 | P75 | Min | Max |
|---|---|---|---|---|---|---|---|---|
| Mercor | 1,260 | 1,230 | $80.00 | $82.17 | $50.00 | $100.00 | $1.00* | $400.00 |
| Alignerr | 34 | 34 | $80.00 | $95.81 | $80.00 | $125.00 | $57.50 | $190.00 |
| SME Careers | 501 | 490 | $55.00 | $56.75 | $30.00 | $75.00 | $3.00* | $175.00 |
| Micro1 | 488 | 455 | $48.50 | $53.87 | $35.00 | $70.00 | $6.00 | $182.00 |
| Turing | 239 | 239 | $27.00 | $32.82 | $21.00 | $47.00 | $15.00 | $65.00 |
| Rex (n<10) | 9 | 8 | $105.00 | $91.50 | $61.25 | $110.00 | $47.00 | $120.00 |
| Vetto (n<10) | 6 | 6 | $45.00 | $65.83 | $40.00 | $83.75 | $40.00 | $170.00 |
Rates normalized to USD/hour. Active listings census as of July 2026. Rex ($n=8$) and Vetto ($n=6$) listed for reference but fall below the $n \ge 10$ threshold for generalized conclusions. *Mercor's $1.00 and SME Careers' $3.00 minimums are outlier listings (piece-rate or partially unparsed pay strings) rather than representative floor offers; they are retained for transparency but should not be read as typical entry pay.
Finding 2: Pay Dispersion and Rate Volatility
The interquartile range (P25 to P75) highlights distinct compensation structures across platforms:
- Mercor: Features the widest rate spread ($50.00/hr to $100.00/hr IQR; maximum $400.00/hr), reflecting a market where generalist annotators start around $20β$30/hr while niche domain specialists (legal, medical, AI safety) command $100β$400/hr.
- Alignerr: Exhibits a high floor and ceiling ($80.00/hr to $125.00/hr IQR), reflecting platform positioning focused almost exclusively on advanced technical and specialized academic roles.
- Micro1 & SME Careers: Show tight, predictable middle-tier bands ($35.00β$70.00/hr and $30.00β$75.00/hr IQRs respectively), providing steady baseline compensation for general domain training and bilingual evaluation.
- Turing: Displays the narrowest overall rate corridor ($21.00/hr to $47.00/hr IQR), indicating highly standardized compensation bands across global software engineering pools.
Limitations
Advertised Rates vs. Realized Earnings. This analysis measures advertised hourly rates from public job postings. It does not account for unpaid onboarding, qualification tests, or platform commission fees.
Geographic Rate Variances. Listings offering geo-specific rates are aggregated as advertised; actual candidate offers may vary based on country location.
Advertised rates are not guaranteed offers. These platforms don't always pay what's advertised to every applicant. Top-of-range rates are typically reserved for candidates in Tier 1 countries (US, UK, Canada, etc.), and actual offers vary by location, experience level, qualifications, and role fit. Treat the figures in this report as the advertised ceiling and floor of the market, not a guarantee of what any individual candidate will be offered.
Realized vs. advertised pay. This report measures rates as advertised in open job postings, not audited payout data. We cannot verify what percentage of listings convert candidates at the advertised max, median, or min rate.
Conclusion
The AI evaluation market has stratified into three distinct compensation tiers: specialized expert platforms (Mercor, Alignerr) leading at $80.00/hr median, mid-tier domain platforms (SME Careers, Micro1) operating in the $48β$55/hr range, and global developer platforms (Turing) centered around $27.00/hr.

Pietro R.
MSc Human-Computer Interaction | Founder & Product Owner
Pietro is the founder and technical lead of aitrainer.work. He builds and maintains the platform's data pipeline, certification infrastructure, and editorial standards.