Internet-Native Bilingual Evaluator Expert (Japanese)
Mercor • Remote, Japan
Education
Not stated
Type
Hourly
Pay Rate
$30–$40/hr
Listed
105d ago
Apply opens Mercor in a new tab.
Apply Now → ⚡ Boost your chances - Optimize your resume with Rezi.aiWhat We Know About This Role
- Weekly hours
- 10–20 hrs/week
About this role
From the Mercor listing
Mercor is seeking Internet-Native Bilingual Evaluator Experts (Japanese–English) to contribute to advanced AI research focused on multilingual communication, internet culture, and contemporary digital language use. This role is ideal for individuals who actively participate in both Japanese- and English-language online communities and understand how internet culture, slang, humour, and social norms differ across languages.
• Analyse content from social media, forums, video platforms, gaming communities, and digital communication channels.
Job Responsibilities
Evaluate Internet-Native Communication
Analyse content from social media, forums, video platforms, gaming communities, and digital communication channels. Assess AI-generated outputs for linguistic quality, cultural relevance, and authenticity. Evaluate internet slang, memes, abbreviations, emoji usage, honorific language, and code-switching.
- Analyse content from social media, forums, video platforms, gaming communities, and digital communication channels.
- Assess AI-generated outputs for linguistic quality, cultural relevance, and authenticity.
- Evaluate internet slang, memes, abbreviations, emoji usage, honorific language, and code-switching.
Develop Evaluation Standards
Create evaluation rubrics for internet-native bilingual communication. Document linguistic and cultural edge cases specific to Japanese and English online communities. Establish standards for tone, appropriateness, and contextual understanding.
- Create evaluation rubrics for internet-native bilingual communication.
- Document linguistic and cultural edge cases specific to Japanese and English online communities.
- Establish standards for tone, appropriateness, and contextual understanding.
Conduct Model Testing and Feedback
Test AI systems using bilingual prompts and real-world internet scenarios. Evaluate outputs in Japanese and English. Provide detailed feedback to improve multilingual model performance.
- Test AI systems using bilingual prompts and real-world internet scenarios.
- Evaluate outputs in Japanese and English.
- Provide detailed feedback to improve multilingual model performance.
Support Quality Assurance
Participate in benchmark development and QA reviews. Ensure consistency and quality across evaluation datasets.
- Participate in benchmark development and QA reviews.
- Ensure consistency and quality across evaluation datasets.
Minimum Qualifications
Native or near-native fluency in Japanese and professional fluency in English. Deep familiarity with Japanese and English internet culture. Active engagement with platforms such as X, YouTube, Discord, Reddit, Twitch, Nico Nico, gaming communities, or similar online ecosystems. Strong writing, analytical, and critical thinking skills. Understanding of internet slang, memes, internet etiquette, and online communication norms in both languages. Available to commit 10–20 hours per week.
- Native or near-native fluency in Japanese and professional fluency in English.
- Deep familiarity with Japanese and English internet culture.
- Active engagement with platforms such as X, YouTube, Discord, Reddit, Twitch, Nico Nico, gaming communities, or similar online ecosystems.
- Strong writing, analytical, and critical thinking skills.
- Understanding of internet slang, memes, internet etiquette, and online communication norms in both languages.
- Available to commit 10–20 hours per week.
Preferred Qualifications
Background in linguistics, translation, localisation, journalism, social media, humanities, or AI evaluation. Experience evaluating user-generated content or online communities. Familiarity with internet subcultures, fandoms, gaming communities, and contemporary digital trends. Interest in AI, language models, and multilingual communication research. We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.
- Background in linguistics, translation, localisation, journalism, social media, humanities, or AI evaluation.
- Experience evaluating user-generated content or online communities.
- Familiarity with internet subcultures, fandoms, gaming communities, and contemporary digital trends.
- Interest in AI, language models, and multilingual communication research.
Requirements
- Must be eligible to work in one of: Remote, Japan
- Fluent proficiency in English (Written & Verbal)
- Reliable high-speed internet connection
How long hiring takes
Across the AI training platforms we refer candidates to, the median gap between referral and hire is about 30 days. It varies by platform and role, so treat it as a rough guide for this one.
Talent Pool members
Apply through this link and we can put you forward to Mercor when your profile is a strong match. Not every applicant is submitted. If you're not in the pool yet, set up your profile first.
Set up your profile →What to Expect
Looking at Mercor Languages listings we've tracked, contracts in this domain typically run about 23.5 weeks. Actual length varies by project, but this gives you a realistic baseline going in.
Based on 64 extracted Mercor Languages listings.
Eligible Languages
Fluent proficiency in English or Japanese
Why this role
At $30–$40/hr, this Internet-Native Bilingual Evaluator Expert (Japanese) position compensates Languages expertise on its own terms, the kind of rate that would otherwise only show up inside a law firm, hospital, or research lab.
Talent pool
We're light on Languages candidates
We've matched 72 people with a Languages background against 925 Languages listings we've tracked, so most go out without one. Set up a profile and we'll consider you for a role like this one.
Set up your profileSkills and categories
Explore other opportunities in related specializations:
Related jobs
Browse All Jobs from Mercor
Discover more opportunities on Mercor that match your skills and interests.
View All Mercor Jobs →Verified Reviews
Community Reviews
Share your experience with Mercor
Help other candidates make better decisions by leaving a review.
Sign in to leave a reviewLeave your review
Common questions
Does it cost money to apply to Mercor?
No, applying and joining Mercor is free. Mercor's revenue comes from a fee it charges the client on top of your hourly rate, not from applicants. Treat any request for payment to join as a red flag.
Is Mercor for freelancers or full-time contractors?
Mercor places you with one client for a defined engagement, like 'Python Tutor for 3 months', rather than having you grab small tasks from a shared queue. Most roles function as steady contract work, not one-off gigs.
What does task-based AI training work look like?
Practical, hands-on data work: recording short videos, categorizing images, rating text responses, or analyzing data. Tasks are designed to be short and distinct, typically 5 to 60 minutes each.
What does asynchronous AI training work mean in practice?
No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.
What does Languages work look like for an Internet-Native Bilingual Evaluator Expert (Japanese)?
Tasks here are scoped to Languages, not generic labeling. As an Internet-Native Bilingual Evaluator Expert (Japanese), expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Languages) rather than following a one-size-fits-all rubric. If you don't have hands-on Languages background, this is likely not the right listing to start with.
What specific skills does this listing call for?
Bilingual, English, Japanese, and Multilingual are named directly in the listing. If you don't have hands-on experience with these, expect the screening process to test for them directly rather than accepting adjacent experience as a substitute.
How many hours per week does this role require?
Based on the listing, this role is scoped at 10–20 hours per week. Treat this as a real commitment expectation, not a loose estimate.
How much does this specific role pay?
This listing is posted at $30–$40/hr, an hourly rate. The range reflects experience level and negotiated terms, not a placeholder, so where you land in it depends on your background and the assessment. Pay can change between when we last checked the listing and when you apply, so confirm the current number on the platform's own application page before committing time.
Do I need to be fluent in English or Japanese?
Yes. This role specifically requires English or Japanese proficiency, on top of the English most AI training work assumes by default. If English is not a language you're fluent in, this is not the right role. Filter for your native language to find better-matched listings.
What happens when I click Apply on this listing?
You'll be taken to Mercor's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.
Can I apply from outside Japan?
This specific role is open only to people based in Japan. If you are somewhere else, applying is unlikely to lead to an offer even if you pass the assessment, because the restriction is usually about where the work can legally be contracted rather than your skills. Read the full description for any tax-residency or right-to-work caveats before you apply.
How soon will I start working after applying to Mercor?
Not immediately. Mercor is a talent marketplace, not a task queue, so applying puts you in a pool of candidates. You start working only once a specific client, like a major AI lab, selects your profile, and that matching process can take weeks.