Skip to content
aitrainer.work - AI Training Jobs Platform
Healthcare mercor

Applied Health & Medicine Benchmark Specialist

Mercor • Remote

Education

PhD

Type

Hourly

Pay Rate

$94–$119/hr

Listed

51d ago

Apply opens Mercor in a new tab.

Apply Now →

What We Know About This Role

Weekly hours
10 hrs/week

About this role

From the Mercor listing

Role Overview

We are seeking expert medical and health science professionals to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core health and medicine domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types:

  • Question Authoring — Create original, challenging multiple-choice questions in your area of medical expertise, rate their difficulty, and submit them for review.

  • Question Verification — Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made.

Health & Medicine Domains Covered

Clinical Medicine & Surgery, Medical Imaging & Diagnostics, Pharmacovigilance, Healthcare Management & Economics, Rehabilitation and Allied Health.

Key Responsibilities

  • Author original health and medicine questions that test deep conceptual understanding, not surface-level recall

  • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement

  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above)

  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers

  • Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format

  • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, clinical guidelines)

  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made

Ideal Qualifications

  • MD, DO, PhD, or doctoral candidate in Medicine, Biomedical Sciences, Public Health, or a closely related field

  • Master's degree considered for candidates with exceptional depth in a specific subdomain

  • Strong command of graduate-level medical knowledge, clinical reasoning, and biomedical research methodology

  • Board certification, clinical experience, or research publications in health fields is a strong plus

  • Excellent written English and ability to express complex ideas clearly and concisely

More About the Opportunity

  • Expected commitment: 10+ hours/week

  • Asynchronous, fully remote work

Requirements

  • Must be eligible to work in Remote
  • Fluent proficiency in English (Written & Verbal)
  • Reliable high-speed internet connection
  • PhD's degree or equivalent professional experience
  • Demonstrated expertise in Healthcare
  • Valid professional license or certification in the relevant clinical/health field (MD, DO, RN, RD, NBHWC, etc.)
  • Current knowledge of medical standards and practices

How long hiring takes

Across the AI training platforms we refer candidates to, the median gap between referral and hire is about 30 days. It varies by platform and role, so treat it as a rough guide for this one.

Within 2 weeks
~25%
Within 6 weeks
~60%
Within 3 months
~80%

Talent Pool members

Apply through this link and we can put you forward to Mercor when your profile is a strong match. Not every applicant is submitted. If you're not in the pool yet, set up your profile first.

Set up your profile →

Why this role

$94–$119/hr for Applied Health & Medicine Benchmark Specialist work covers both income and flexibility. You set your own hours, and the work draws on Healthcare knowledge you already have instead of asking you to learn a new field.

Talent pool

We're light on Healthcare candidates

We've matched 33 people with a Healthcare background against 373 Healthcare listings we've tracked, so most go out without one. Set up a profile and we'll consider you for a role like this one.

Set up your profile

Skills and categories

Explore other opportunities in related specializations:

Related jobs

Mercor

Browse All Jobs from Mercor

Discover more opportunities on Mercor that match your skills and interests.

View All Mercor Jobs →

Verified Reviews

Loading reviews…

Community Reviews

Loading reviews…
💬

Share your experience with Mercor

Help other candidates make better decisions by leaving a review.

Sign in to leave a review

Common questions

Is Mercor for freelancers or full-time contractors?

Mercor places you with one client for a defined engagement, like 'Python Tutor for 3 months', rather than having you grab small tasks from a shared queue. Most roles function as steady contract work, not one-off gigs.

Does Mercor's application require an on-camera interview?

Yes, every applicant records a video interview with an AI interviewer that asks questions about your resume. Clients review that recording to judge communication skills before matching, so there's no way to apply without going on camera.

Why do these AI training roles pay so much?

Because general knowledge isn't what's being tested. The model already knows the basics; what it needs is expertise on edge cases, the rare, difficult, highly technical judgment calls only a senior professional in the field would make correctly.

What does the day-to-day workload look like for elite-expert AI training roles?

Slow and deep, not fast and repetitive. A single task can take 45-60 minutes of researching citations or verifying complex calculations. Quality is what's being measured here, not throughput.

What does Healthcare work look like for an Applied Health & Medicine Benchmark Specialist?

Tasks here are scoped to Healthcare, not generic labeling. As an Applied Health & Medicine Benchmark Specialist, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Healthcare) rather than following a one-size-fits-all rubric. If you don't have hands-on Healthcare background, this is likely not the right listing to start with.

What specific skills does this listing call for?

English and Expert are named directly in the listing. If you don't have hands-on experience with these, expect the screening process to test for them directly rather than accepting adjacent experience as a substitute.

How many hours per week does this role require?

Based on the listing, this role is scoped at about 10 hours per week. Treat this as a real commitment expectation, not a loose estimate.

How much does this specific role pay?

This listing is posted at $94–$119/hr, an hourly rate. The range reflects experience level and negotiated terms, not a placeholder, so where you land in it depends on your background and the assessment. Pay can change between when we last checked the listing and when you apply, so confirm the current number on the platform's own application page before committing time.

What happens when I click Apply on this listing?

You'll be taken to Mercor's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.

Is a PhD required?

For this specific role, yes, or near-equivalent professional depth. The credential gate is enforced at the assessment stage, not just on paper. That said, active PhD candidates and people with equivalent published research have qualified without a formal degree. The assessment is the real filter.

How soon will I start working after applying to Mercor?

Not immediately. Mercor is a talent marketplace, not a task queue, so applying puts you in a pool of candidates. You start working only once a specific client, like a major AI lab, selects your profile, and that matching process can take weeks.