Skip to content
aitrainer.work - AI Training Jobs Platform
Psychology mercor

AI Safety Experts — English & Gujarati

🇮🇳

Mercor Remote, India

Education

Any

Type

hourly

Pay Rate (by country)

$20–$22/hr

Listed

55d ago

✅ Applying through this link supports our platform at no cost to you.

This position is hosted on an external talent platform. Please only apply for this position if it fits your skills and interests.

Apply Now

In our Talent Pool?

Apply through this link and we can vouch for you to Mercor. ? We vouch for Talent Pool members who apply through our referral link, when we believe they're a strong match. Not every applicant gets a vouch. Not in the pool yet? Set up your profile first.

Set up your profile →

Mercor: our referral track record

We've referred 191 candidates to Mercor roles. 14% (27) were placed.

About this Role

From the Mercor listing

Location: Remote

Fluent Language Skills Required: English & Gujarati. Native fluency in English and Gujarati is required for this position.

Why This Role Exists

At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who probe AI models with adversarial inputs, surface vulnerabilities, and generate the red team data that makes AI safer for our customers.

This project involves reviewing AI outputs that touch on sensitive topics such as bias, misinformation, or harmful behaviors. All work is text-based, and participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources. Before being exposed to any content, the topics will be clearly communicated.

What You’ll Do

  • Red team conversational AI models and agents: jailbreaks, prompt injections, misuse cases, bias exploitation, multi-turn manipulation

  • Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks

  • Apply structure: follow taxonomies, benchmarks, and playbooks to keep testing consistent

  • Document reproducibly: produce reports, datasets, and attack cases customers can act on

Who You Are

  • You bring prior red teaming experience (AI adversarial work, cybersecurity, socio-technical probing)

  • You’re curious and adversarial: you instinctively push systems to breaking points

  • You’re structured: you use frameworks or benchmarks, not just random hacks

  • You’re communicative: you explain risks clearly to technical and non-technical stakeholders

  • You’re adaptable: thrive on moving across projects and customers

Nice-to-Have Specialties

  • Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction

  • Cybersecurity: penetration testing, exploit development, reverse engineering

  • Socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing

  • Creative probing: psychology, acting, writing for unconventional adversarial thinking

What Success Looks Like

  • You uncover vulnerabilities automated tests miss

  • You deliver reproducible artifacts that strengthen customer AI systems

  • Evaluation coverage expands: more scenarios tested, fewer surprises in production

  • Mercor customers trust the safety of their AI because you’ve already probed it like an adversary

Why Join Mercor

  • Build experience in human data-driven AI red teaming at the frontier of safety

  • Play a direct role in making AI systems more robust, safe, and trustworthy

Requirements

  • Must be eligible to work in one of: Remote, India
  • Fluent proficiency in English (Written & Verbal)
  • Reliable high-speed internet connection

Eligible Languages

Fluent proficiency in English or Gujarati

English Gujarati

Why This Role

Rare opportunity for top 1% experts. Earn $21/hr contributing to the world's most advanced AI labs. This is one of the few roles where academic precision is valued as highly as commercial output.

Skills & Categories

Explore other opportunities in related specializations:

Related Jobs

Mercor

Browse All Jobs from Mercor

Discover more opportunities on Mercor that match your skills and interests.

View All Mercor Jobs →

Verified Reviews

Loading reviews…

Community Reviews

Loading reviews…
💬

Share your experience with Mercor

Help other candidates make better decisions by leaving a review.

Sign in to leave a review

Frequently Asked Questions

Is Mercor for freelancers or full-time contractors?

Mercor places you with one client for a defined engagement, like 'Python Tutor for 3 months', rather than having you grab small tasks from a shared queue. Most roles function as steady contract work, not one-off gigs.

Does Mercor's application require an on-camera interview?

Yes, every applicant records a video interview with an AI interviewer that asks questions about your resume. Clients review that recording to judge communication skills before matching, so there's no way to apply without going on camera.

Does it cost money to apply to Mercor?

No, applying and joining Mercor is free. Mercor's revenue comes from a fee it charges the client on top of your hourly rate, not from applicants. Treat any request for payment to join as a red flag.

What does task-based AI training work actually look like?

Practical, hands-on data work: recording short videos, categorizing images, rating text responses, or analyzing data. Tasks are designed to be short and distinct, typically 5 to 60 minutes each.

What does asynchronous AI training work mean in practice?

No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.

What does Psychology work look like for a AI Safety Experts — English & Gujarati?

Tasks here are scoped to Psychology, not generic labeling. As a AI Safety Experts — English & Gujarati, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Psychology) rather than following a one-size-fits-all rubric. If you don't have hands-on Psychology background, this is likely not the right listing to start with.

Do I need to be fluent in English or Gujarati?

Yes. This role specifically requires English or Gujarati proficiency. You will likely be evaluated on written fluency during the assessment, not just conversational level. If English is not your first language or you are not professionally fluent, this is not the right role. Filter for your native language to find better-matched listings.

What happens when I click Apply on this listing?

You'll be taken to Mercor's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.

Can I apply from outside India?

This specific role is restricted to India. If you are outside these locations, applying is unlikely to result in an offer even if you pass the assessment. Check the full job description for any VPN or tax-residency caveats.

How soon will I start working after applying to Mercor?

Not immediately. Mercor is a talent marketplace, not a task queue, so applying puts you in a pool of candidates. You start working only once a specific client, like a major AI lab, selects your profile, and that matching process can take weeks.