Generalist Expert | $70/hr Remote
Overview
This role involves assessing AI-generated responses and offering detailed written evaluations to guide model improvement. You'll apply sharp analytical skills to identify subtle reasoning flaws and articulate clear, evidence-based feedback. It's a chance for strong critical thinkers to directly influence cutting-edge AI research projects.
What You'll Do5
- 1Review AI-generated outputs for accuracy, coherence, and subtle reasoning gaps, then write structured evaluations.
- 2Compose clear, well-supported written rationales that go beyond surface-level observations.
- 3Apply detailed evaluation guidelines consistently and honestly, including giving critical assessments when needed.
- 4Work independently to complete assignments while strictly following task instructions and quality standards.
- 5Ensure all feedback is produced without the use of AI writing assistance tools.
Requirements5
- 1A bachelor's degree from a globally top-ranked university is preferred, but strong analytical skills matter most.
- 2Exceptional critical reading ability to detect nuance, implicit meaning, and logical inconsistencies.
- 3Excellent written communication skills for producing precise, evidence-based rationales.
- 4Native-level English fluency (C2 proficiency) is required.
- 5Ability to work autonomously with strict attention to detail and honest, consistent judgment.
Who Should Apply
We're looking for sharp, detail-oriented professionals who enjoy deep analysis and providing constructive, honest feedback. If you have a knack for spotting subtle errors in reasoning and can articulate why something is right or wrong in a clear, structured way, this role will let you contribute to improving AI systems. The ideal candidate works well independently and takes pride in maintaining high standards without relying on AI tools.
Salary Insight
This is a remote, hourly role paying $70/hr.
Application Tip
To stand out, prepare a short (1-2 paragraph) critique of a recent AI-generated text sample — showing how you would evaluate its reasoning, clarity, and completeness. Attach it as a writing sample to demonstrate your analytical and communication style.
Similar open positions
Explore active roles that match your skills and interests.
Micro1
VerifiedAI Evaluation Specialist | $20-$35/hr Remote
As an AI Evaluation Specialist, you'll help train next-generation AI systems by designing and executing hands-on evaluation tasks. Your insights will directly shape how models learn, reason, and perform on practical computer-based workflows. This is a fully remote contract role where meticulous observation and clear documentation are key.
Micro1
VerifiedImage Evaluation Generalist | $20-$30/hr Remote
Micro1 is looking for sharp-eyed contractors to join a remote team evaluating images that will help train the next wave of AI systems. In this role, you'll apply your own expertise to spot inconsistencies and provide clear feedback — no AI background needed. Your assessments will directly influence how models learn to interpret visual data.
Mercor
VerifiedDesign Expert | $80-$150/hr Remote
This role offers experienced designers a chance to influence how leading AI systems approach creative work. Instead of producing designs, you’ll set the bar for quality by crafting scoring criteria and evaluating both AI-generated and human-created samples. Your detailed written feedback will directly sharpen the AI’s design judgment.
Micro1
VerifiedAI Evaluation Analyst | $20-$30/hr Remote
As an AI Evaluation Analyst at micro1, you'll help train advanced language models by creating high-quality evaluation data and multi-turn conversations. This remote contract role is ideal for someone with strong analytical and writing skills who wants to directly influence how AI systems reason and interact. No previous AI experience is required—your domain expertise and attention to detail are what matter.
Mercor
VerifiedMarket research / competitive intelligence Evaluator | $80-$120/hr Remote
We're seeking an experienced market research and competitive intelligence professional to evaluate AI-generated work products. You'll review documents, spreadsheets, and slide decks for accuracy and domain quality, using your expertise to grade outputs. This is a remote, contract role with a flexible hourly schedule.
Mercor
VerifiedGeneralist - English & Tamil | $15-$20/hr Remote
This remote contract role is perfect for a bilingual Tamil and English speaker who enjoys digging into the details of AI-generated content. You'll evaluate model responses in Tamil, checking for accuracy, clarity, and reasoning quality, then provide structured feedback in English. Your work will directly help shape the next generation of conversational AI, making it smarter and more reliable.