AI Safety Practitioner | $60-$70/hr Remote
Overview
We’re seeking a sharp evaluator to test frontier AI models for safety, quality, and alignment. You’ll assess responses on complex, policy-sensitive topics like misinformation and biosecurity, applying structured rubrics to catch unsafe or misaligned outputs. Your feedback will directly shape how these models behave for millions of users worldwide.
What You'll Do6
- 1Analyze AI-generated responses for factual accuracy, safety compliance, and policy adherence across high-stakes domains.
- 2Review content involving sensitive areas such as misinformation, political persuasion, self-harm, violence, cyber threats, and biosecurity.
- 3Apply and continuously refine evaluation rubrics used in RLHF, SFT, and AI safety benchmarking.
- 4Identify unsafe outputs, hallucinations, reasoning failures, and policy violations in model responses.
- 5Provide structured, actionable feedback to improve model alignment and safety performance.
- 6Collaborate with AI researchers and safety teams on ongoing evaluation projects.
Requirements4
- 1A Bachelor’s degree or higher in Journalism, Communications, Psychology, Sociology, Public Policy, Law, Biology, Chemistry, Computer Science, or a related field.
- 25+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related discipline.
- 3Excellent written English, critical thinking, and analytical reasoning skills.
- 4Consistent judgment when evaluating nuanced, policy-sensitive, or ambiguous scenarios.
Who Should Apply
This role is perfect for someone with a background in journalism, policy, science, or research who enjoys dissecting complex, grey-area content. You’re ethically driven, detail-oriented, and passionate about responsible AI. Experience with RLHF, content moderation, or developing safety rubrics is a plus, but a sharp analytical mind matters most.
Salary Insight
The role pays $60.00–$70.00 per hour on a remote, contract basis.
Application Tip
In your cover letter, describe a specific instance where you evaluated a nuanced ethical or safety issue and influenced a decision. Show that you can handle grey-area scenarios with clear reasoning.
Similar open positions
Explore active roles that match your skills and interests.
Mercor
VerifiedAI Safety Red Teamer | $70-$84/hr Remote
Join a team focused on strengthening the safety of advanced AI systems by uncovering their hidden weaknesses. As an AI Safety Red Teamer, you'll design clever prompts to stress-test models, spot dangerous behaviors, and help improve how these systems handle tricky real-world scenarios. This remote contractor role lets you work alongside top researchers while earning $70–$84 per hour.
Mercor
VerifiedCybersecurity SWE — AI Safety | $60-$90/hr Remote
This part-time remote role invites a cybersecurity engineer to apply their skills in AI safety — no prior machine learning background needed. You'll work hands-on evaluating how advanced AI models handle sensitive security topics, with full training on the workflow. It's a unique chance to blend deep security knowledge with the frontier of AI development.
Micro1
VerifiedAI Content Evaluation Specialist | $25-$35/hr Remote
This role puts your content moderation expertise to work in training the next generation of AI systems. You'll evaluate and refine how models handle images and text, ensuring they align with safety standards. Your domain knowledge matters more than AI experience—if you understand digital platforms and content risks, you're a strong fit.
Mercor
VerifiedAI Safety Experts — English & Assamese | $20-$22/hr Remote
We're assembling a remote team of AI red teamers to stress-test conversational models and uncover hidden vulnerabilities. In this role, you'll probe AI systems with adversarial inputs—jailbreaks, prompt injections, bias exploitation—and generate the high-quality human data that makes AI safer for real-world use. Fluency in both English and Assamese is essential, as you'll review sensitive outputs and document reproducible attack cases. This project is ideal for structured, curious thinkers who enjoy pushing systems to their limits.
Mercor
VerifiedBiology Expert (PhD) — AI Safety | $65-$70/hr Remote
We’re seeking PhD-level biologists to lend their expertise to an important mission: making advanced AI systems safer and more reliable. You’ll evaluate how well these models handle complex life‑science topics, from molecular biology to immunology, and help refine their responses. No previous AI experience is required — training on the evaluation workflow is provided.
Micro1
VerifiedChild & Online Safety Expert | $50-$90/hr Remote
micro1 is seeking experienced Child & Online Safety experts to help train next-generation AI systems focused on protecting young people online. You'll apply your deep domain knowledge—no AI background required—to build mental-health safety evaluation frameworks and digital risk-assessment protocols. This contract role lets you directly shape how AI models handle crisis scenarios involving minors.