AI Safety & Policy Analyst
Turing • Remote, Bangladesh, Brazil, Egypt, India, Kenya, Mexico, Nigeria, Pakistan, Turkey, United States
Education
Not stated
Type
Hourly
Listed
186d ago
Apply opens Turing in a new tab.
Apply Now → ⚡ Boost your chances - Optimize your resume with Rezi.aiAbout this role
From the Turing listing
About Turing
Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in coding, reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L.
Role Overview
As an AI Safety & Policy Analyst, you will be on the front lines of developing safe and responsible AI. You will be responsible for challenging our models' safeguards, identifying new vulnerabilities, and creating the detailed evaluation rubrics used to train and test our next generation of large language models. This role requires a unique blend of creativity, analytical rigor, and a deep understanding of policy. You will not just follow instructions; you will actively design the tests, using an adversarial mindset to discover how models fail. You will then use your analytical skills to articulate why they failed, creating the precise rubrics and rationales that teach our models to be safer and more helpful.
*NOTE: This role may involve reviewing or encountering disturbing, sensitive, or otherwise potentially distressing content as part of AI safety evaluations. Candidates selected for this position may be required to sign an acknowledgment form confirming their understanding and consent.
What does day-to-day look like
In this role, you will be part of a dynamic team focused on LLM safety and alignment. Your day-to-day work will involve:
- Designing and executing creative, multi-turn conversational prompts that test model compliance with complex safety policies (e.g., Discriminatory, Abetting, Copyrighted Content, Harmful Advice).
- Identifying, analyzing, and documenting model failures, including successful jailbreaks and subtle policy violations.
- Developing detailed, objective, and independent rubrics for new safety prompts, assigning priority scores (e.g., Crucial, Important, Less Important) to define and weight desired model behavior.
- Rigorously evaluating and stack-ranking multiple model responses to a single prompt, using the rubrics you created to ensure clear discrimination between good, bad, and nuanced failures.
- Writing clear, defensible "Single Rationales" for your rankings that explain the "why" behind your evaluation, focusing on both safety and quality.
- Collaborating with researchers and policy-makers to understand new risks and refine the safety taxonomy.
Education & Experience
- BS/BA degree or equivalent experience in a relevant field (e.g., Policy, Law, Ethics, Linguistics, Journalism, Computer Science, or a related analytical field).
- Experience in content moderation, policy analysis, AI safety evaluation, or a related role is strongly preferred
Requirements
- English Proficiency: Ability to read and write in English with a high degree of comp.
- Exceptional Analytical Thinking: A proven ability to research and evaluate nuanced, complex, and ambiguous information against a defined set of policy criteria.
- Creative & Adversarial Mindset: Experience in "red teaming," prompt engineering, or designing creative challenge prompts intended to test and bypass AI safety filters.
- Strong Policy & Taxonomy Acumen: A strong understanding of Trust & Safety principles, particularly in relation to LLMs (e.g., categories like misinformation, abetting, bias/stereotypes, jailbreaks, and dual-use). We welcome candidates with expertise in at least one of the following domains: Cyberharm Violence, terrorism Bias and stereotypes Mental health and self-harm Child safety Nudity and sexually explicit content Misinformation Fraud Sycophancy Regulated goods Privacy and identity rights Copyright Legal, medical, financial information
- Meticulous Attention to Detail: The ability to design and author precise, self-contained, and independent evaluation rubrics that can clearly discriminate between models.
- Excellent Written Communication: Superior ability to articulate complex rationale for model rankings clearly and concisely, providing a strong training signal for engineers.
- Familiarity with RLHF (Reinforcement Learning from Human Feedback) workflows and data annotation is a significant plus.
- Feedback: Ability to provide constructive feedback and detailed annotations.
- Communication: Excellent communication and collaboration skills.
- Independence: Self-motivated and able to work independently in a remote setting.
- Technical Setup: Desktop/Laptop set up with a good internet connection.
Benefits
- Flexible working hours and remote work environment.
- Opportunity to work on cutting-edge AI projects with leading LLM companies.
- Potential for contract extension based on performance and project needs.
Offer Details
- Commitments Required : at least 4 hours per day and a total of 40 hours per week with 2-4 hours of overlap with PST.
- Engagement type : Contractor assignment/freelancer (no medical/paid leave)
- Duration of contract: 1 month
- This role will require some overlap with UTC-8:00 (2-4 hrs/day) America/Los_Angeles
Application Process
- Shortlisted candidates will be sent automated analytical challenges.
- Once you clear them, you are ready to go!
How long hiring takes
Across the AI training platforms we refer candidates to, the median gap between referral and hire is about 30 days. It varies by platform and role, so treat it as a rough guide for this one.
What to Expect
Looking at Turing Software Engineering listings we've tracked, contracts in this domain typically run about 8.4 weeks. Actual length varies by project, but this gives you a realistic baseline going in.
Based on 28 extracted Turing Software Engineering listings.
Why this role
Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who speci
Skills and categories
Explore other opportunities in related specializations:
Related jobs
Browse All Jobs from Turing
Discover more opportunities on Turing that match your skills and interests.
View All Turing Jobs →Verified Reviews
Community Reviews
Share your experience with Turing
Help other candidates make better decisions by leaving a review.
Sign in to leave a reviewLeave your review
Common questions
Do I need to be a software engineer to work for Turing?
No, not anymore. Turing built its name matching senior engineers with Silicon Valley companies, but it has since expanded into AGI infrastructure work and now hires non-engineering domain experts, technical writers, and researchers for post-training data annotation and RLHF. A strong analytical background and excellent English matter more than coding ability.
How does Turing's talent matching work?
Turing calls it the Intelligent Talent Cloud. You build a profile and go through vetting (automated tests, an AI-powered interview, practical skill assessments), and once vetted, Turing's algorithm surfaces your profile directly to partner companies like Fortune 500s and top AI labs. You don't browse listings or bid on work; matches come to you.
What does asynchronous AI training work mean in practice?
No set hours, no check-ins, no meetings. You log in when you want, pick up an available task, complete it, and submit; nobody is waiting on you in real time. That's different from remote employment, where you're expected online during business hours. The tradeoff: you're competing with others for available tasks, so an empty queue means there's simply nothing to do until more work is released.
What does Software Engineering work look like for an AI Safety & Policy Analyst?
Tasks here are scoped to Software Engineering, not generic labeling. As an AI Safety & Policy Analyst, expect to draw on real domain judgment (evaluating outputs, correcting errors, or providing expert reasoning specific to Software Engineering) rather than following a one-size-fits-all rubric. If you don't have hands-on Software Engineering background, this is likely not the right listing to start with.
What specific skills does this listing call for?
English is named directly in the listing. If you don't have hands-on experience with this, expect the screening process to test for it directly rather than accepting adjacent experience as a substitute.
How much does this specific role pay?
The listing doesn't state a rate. The $25–$55/hr shown here is our estimate from the role type and location (see /pay-methodology), so treat it as a rough guide and confirm the actual rate with the platform before committing time.
What happens when I click Apply on this listing?
You'll be taken to Turing's external site to complete your application there. This listing links through a referral, but the process is identical to applying directly; the link just routes you correctly. Create an account on their site and follow their onboarding steps.
Can I apply from outside Bangladesh, Brazil, Egypt and 7 other countries?
This specific role is open only to people based in Bangladesh, Brazil, Egypt, India, Kenya, Mexico, Nigeria, Pakistan, Turkey, and the United States. If you are somewhere else, applying is unlikely to lead to an offer even if you pass the assessment, because the restriction is usually about where the work can legally be contracted rather than your skills. Read the full description for any tax-residency or right-to-work caveats before you apply, since they can differ by country.