
AI Safety Practitioner | $60-$70/hr Remote
Listing checked July 17, 2026 · pay as published by Mercor
Overview
We’re seeking a sharp evaluator to test frontier AI models for safety, quality, and alignment. You’ll assess responses on complex, policy-sensitive topics like misinformation and biosecurity, applying structured rubrics to catch unsafe or misaligned outputs. Your feedback will directly shape how these models behave for millions of users worldwide.
What You'll Do6
- 1Analyze AI-generated responses for factual accuracy, safety compliance, and policy adherence across high-stakes domains.
- 2Review content involving sensitive areas such as misinformation, political persuasion, self-harm, violence, cyber threats, and biosecurity.
- 3Apply and continuously refine evaluation rubrics used in RLHF, SFT, and AI safety benchmarking.
- 4Identify unsafe outputs, hallucinations, reasoning failures, and policy violations in model responses.
- 5Provide structured, actionable feedback to improve model alignment and safety performance.
- 6Collaborate with AI researchers and safety teams on ongoing evaluation projects.
Requirements4
- 1A Bachelor’s degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
- 25+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related discipline.
- 3Excellent written English, critical thinking, and analytical reasoning skills.
- 4Consistent judgment when evaluating nuanced, policy-sensitive, or ambiguous scenarios.
Who Should Apply
This role is perfect for someone with a background in journalism, policy, science, or research who enjoys dissecting complex, grey-area content. You’re ethically driven, detail-oriented, and passionate about responsible AI. Experience with RLHF, content moderation, or developing safety rubrics is a plus, but a sharp analytical mind matters most.
Salary Insight
The role pays $60.00–$70.00 per hour on a remote, contract basis.
Pay and demand on NearSkill
Live dataMedian hourly pay, USD
$70/hour
Among 613 live similar roles that publish pay
- Live similar roles
- 878
- Listed in last 30 days
- 426
- Remote
- 97%
Hiring most right now: micro1 (444) · SME Careers (163) · Mercor (122)
Most requested skills
- llm evaluation11%
- ai training10%
- trainer feedback9%
- documentation9%
Figures from Machine Learning & AI roles live on NearSkill when this page loaded. Only listings that publish USD pay are counted. A role can close before you apply, so check the listing itself.
Compare your resume against these live rolesLocation
Application Tip
In your cover letter, describe a specific instance where you evaluated a nuanced ethical or safety issue and influenced a decision. Show that you can handle grey-area scenarios with clear reasoning.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Micro1
VerifiedAI Content Evaluation Specialist
This role puts your content moderation expertise to work in training the next generation of AI systems. You'll evaluate and refine how models handle images and text, ensuring they align with safety standards. Your domain knowledge matters more than AI experience—if you understand digital platforms and content risks, you're a strong fit.

Mercor
VerifiedAI Safety Red Teamer
Join a team focused on strengthening the safety of advanced AI systems by uncovering their hidden weaknesses. As an AI Safety Red Teamer, you'll design clever prompts to stress-test models, spot dangerous behaviors, and help improve how these systems handle tricky real-world scenarios. This remote contractor role lets you work alongside top researchers while earning $70–$84 per hour.

Mercor
VerifiedAI Safety Experts — English & Swedish
This role involves stress-testing AI systems by simulating adversarial attacks in English and Swedish. You'll generate critical safety data by probing models for vulnerabilities like bias, misinformation, and security flaws. As part of a human red team, you'll help ensure AI behaves safely before it reaches users.

SME Careers
VerifiedRed-Teaming QA Lead for AI Safety and Evaluation
As a remote Red-Teaming QA Lead, you guide quality and consistency across AI red-teaming and safety-evaluation work performed by a distributed contractor team. You review ai red-teaming outputs, assess adversarial prompts and risk classifications, and provide precise written feedback to keep guidelines aligned. You’ll maintain project rubrics, coordinate updates for trainers and QAs, and manage documentation across a fast-moving remote workflow using Discord, Google Sheets, and dashboards. This hourly contract role supports SME Careers’ AI data services and helps improve safety training data for leading models.

Mercor
VerifiedBilingual Italian-English AI Safety Evaluator
Help advance AI safety by evaluating how models respond to sensitive topics in Italian. Your linguistic and cultural judgment will shape which prompts seem adversarial and which conversations escalate. Since no AI or ML expertise is needed, we'll train you on our guidelines and workflow. This remote, part-time role suits someone with excellent written reasoning who can document their decisions clearly. While Western Europe is preferred, we'll consider applicants from other regions.

Micro1
VerifiedChild & Online Safety Expert
micro1 is seeking experienced Child & Online Safety experts to help train next-generation AI systems focused on protecting young people online. You'll apply your deep domain knowledge—no AI background required—to build mental-health safety evaluation frameworks and digital risk-assessment protocols. This contract role lets you directly shape how AI models handle crisis scenarios involving minors.

