AI Safety Experts — English & Urdu
Mercor • Remote, India
Education
Any
Type
hourly
Pay Rate (by country)
$20–$22/hr
Listed
55d ago
✅ Applying through this link supports our platform at no cost to you.
This position is hosted on an external talent platform. Please only apply for this position if it fits your skills and interests.
In our Talent Pool?
Apply through this link and we can vouch for you to Mercor. ? We vouch for Talent Pool members who apply through our referral link, when we believe they're a strong match. Not every applicant gets a vouch. Not in the pool yet? Set up your profile first.
Set up your profile →Mercor: our referral track record
We've referred 191 candidates to Mercor roles. 14% (27) were placed.
About this Role
From the Mercor listing
Location: Remote
Fluent Language Skills Required: English & Urdu. Native fluency in English and Urdu is required for this position.
Why This Role Exists
At Mercor, we believe the safest AI is the one that’s already been attacked — by us. We are assembling a red team for this project - human data experts who probe AI models with adversarial inputs, surface vulnerabilities, and generate the red team data that makes AI safer for our customers.
This project involves reviewing AI outputs that touch on sensitive topics such as bias, misinformation, or harmful behaviors. All work is text-based, and participation in higher-sensitivity projects is optional and supported by clear guidelines and wellness resources. Before being exposed to any content, the topics will be clearly communicated.
What You’ll Do
Red team conversational AI models and agents: jailbreaks, prompt injections, misuse cases, bias exploitation, multi-turn manipulation
Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks
Apply structure: follow taxonomies, benchmarks, and playbooks to keep testing consistent
Document reproducibly: produce reports, datasets, and attack cases customers can act on
Who You Are
You bring prior red teaming experience (AI adversarial work, cybersecurity, socio-technical probing)
You’re curious and adversarial: you instinctively push systems to breaking points
You’re structured: you use frameworks or benchmarks, not just random hacks
You’re communicative: you explain risks clearly to technical and non-technical stakeholders
You’re adaptable: thrive on moving across projects and customers
Nice-to-Have Specialties
Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction
Cybersecurity: penetration testing, exploit development, reverse engineering
Socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing
Creative probing: psychology, acting, writing for unconventional adversarial thinking
What Success Looks Like
You uncover vulnerabilities automated tests miss
You deliver reproducible artifacts that strengthen customer AI systems
Evaluation coverage expands: more scenarios tested, fewer surprises in production
Mercor customers trust the safety of their AI because you’ve already probed it like an adversary
Why Join Mercor
Build experience in human data-driven AI red teaming at the frontier of safety
Play a direct role in making AI systems more robust, safe, and trustworthy
Requirements
- Must be eligible to work in one of: Remote, India
- Fluent proficiency in English (Written & Verbal)
- Reliable high-speed internet connection
Eligible Languages
Fluent proficiency in English or Urdu
Why This Role
Rare opportunity for top 1% experts. Earn $21/hr contributing to the world's most advanced AI labs. This is one of the few roles where academic precision is valued as highly as commercial output.
Skills & Categories
Explore other opportunities in related specializations:
Related Jobs
Browse All Jobs from Mercor
Discover more opportunities on Mercor that match your skills and interests.
View All Mercor Jobs →Verified Reviews
Community Reviews
Share your experience with Mercor
Help other candidates make better decisions by leaving a review.
Sign in to leave a reviewLeave your review
Frequently Asked Questions
Is Mercor for freelancers or full-time contractors?
Mercor places you with one client for a defined engagement, like 'Python Tutor for 3 months', rather than having you grab small tasks from a shared queue. Most roles function as steady contract work, not one-off gigs.
Does Mercor's application require an on-camera interview?
Yes, every applicant records a video interview with an AI interviewer that asks questions about your resume. Clients review that recording to judge communication skills before matching, so there's no way to apply without going on camera.
Does it cost money to apply to Mercor?
No, applying and joining Mercor is free. Mercor's revenue comes from a fee it charges the client on top of your hourly rate, not from applicants. Treat any request for payment to join as a red flag.
What does task-based AI training work actually look like?
Practical, hands-on data work: recording short videos, categorizing images, rating text responses, or analyzing data. Tasks are designed to be short and distinct, typically 5 to 60 minutes each.
What does asynchronous AI training work mean in practice?
No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.
What does Languages work look like for a AI Safety Experts — English & Urdu?
Tasks here are scoped to Languages, not generic labeling. As a AI Safety Experts — English & Urdu, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Languages) rather than following a one-size-fits-all rubric. If you don't have hands-on Languages background, this is likely not the right listing to start with.
Do I need to be fluent in English or Urdu?
Yes. This role specifically requires English or Urdu proficiency. You will likely be evaluated on written fluency during the assessment, not just conversational level. If English is not your first language or you are not professionally fluent, this is not the right role. Filter for your native language to find better-matched listings.
What happens when I click Apply on this listing?
You'll be taken to Mercor's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.
Can I apply from outside India?
This specific role is restricted to India. If you are outside these locations, applying is unlikely to result in an offer even if you pass the assessment. Check the full job description for any VPN or tax-residency caveats.
How soon will I start working after applying to Mercor?
Not immediately. Mercor is a talent marketplace, not a task queue, so applying puts you in a pool of candidates. You start working only once a specific client, like a major AI lab, selects your profile, and that matching process can take weeks.