Micro1
Micro1Verified listing
RemoteHot

Social Science Research Assistant for AI Benchmarking

30–50/hr
Remote
Posted August 14, 2026
contract
10 openings

Overview

Remote contractor role supporting a frontier AI benchmarking project. You’ll design and run authentic evaluation tasks drawn from social science methods, then produce realistic research artifacts and grading rubrics. Your domain knowledge guides how models should be trained to reason and perform, with emphasis on rigor and transparent documentation. Prior AI experience isn’t required; focus is on solid social science expertise and careful task construction. Key tools and areas include survey design, qualitative coding, literature reviews, and statistical outputs, all built to reflect real-world research complexity. Survey design, Qualitative coding, Statistical analysis are central to the work.

What You'll Do6

  • 1Design and author advanced evaluation tasks that mirror real-world research in your field, such as surveys, coding schemes, literature assessments, and statistical analyses.
  • 2Source and assemble authentic research artifacts, including survey instruments, coded datasets, outputs from statistics packages, and supporting documents.
  • 3Choose appropriate methods and provide interpretations that align with social science standards and justification.
  • 4Create comprehensive grading rubrics of 35+ items to evaluate methodological choices, execution quality, and interpretive accuracy.
  • 5Ensure tasks capture the nuance and complexity of day-to-day research rather than simplified, classroom-style exercises.
  • 6Collaborate asynchronously with project leads to refine tasks, data, and evaluation criteria while maintaining rigorous documentation.

Requirements7

  • 1Bachelor’s or Master’s degree in sociology, economics, psychology, political science, or a related field
  • 2At least 2 years of hands-on experience with survey design, qualitative coding, literature reviews, and statistical software (e.g., SPSS, R, Stata)
  • 3Proven ability to synthesize diverse sources into well-documented data artifacts
  • 4Strong emphasis on methodological rigor and reproducible research practices
  • 5Experience developing or applying detailed grading rubrics or evaluation guidelines
  • 6Comfort with detailed, methodical tasks across multiple data types and sources
  • 7Excellent written English for preparing technical research content and evaluation materials

Who Should Apply

The ideal candidate is a social science professional with hands-on survey design, coding, and literature review experience who can craft authentic, complex research tasks. This role may not be a fit for those lacking experience with statistical tools or those who prefer highly routine work. Common fit signals include a track record of rigorous, well-documented research artifacts and the ability to justify methodological choices; candidates who struggle with creating extended rubrics or who avoid nuanced data tasks may score lower. The work demands careful attention to detail and a comfort with asynchronous collaboration.

Salary Insight

Pay is task-based and paid per completed item that meets project specifications; no hourly rate is stated. Pay discussions occur later in the process.

Location

TypeRemote
LocationRemote
This is a remote position

Required Skills

Methodological rigorSource synthesisRubric fidelityResearch mastery

Application Tip

Highlight a specific project where you designed a survey or coding scheme and attached a 35+ item rubric to demonstrate your methodological rigor and rubric-building capability.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Micro1

Micro1

21d agoRemotecontract
Hot

Life Sciences Research Assistant for AI Benchmark Development

Remote Life Sciences Research Assistant helping build next generation AI benchmarks. You’ll design authentic evaluation tasks in biology, synthesize sources like raw data and lab protocols, and define clear methods and interpretations for each task. You’ll create multi-item rubrics that assess rigor, data interpretation, and source fidelity, ensuring the work reflects real-world scientific nuance. Collaboration with project coordinators happens asynchronously to refine deliverables. Bachelor’s or Master’s in biology or related field and hands-on lab experience are preferred.

30–50/hr
· 10 openings
Methodological RigorSource SynthesisRubric Fidelity+2 more
Micro1

Micro1

21d agoRemotecontract
Hot

Physical Sciences Research Assistant Focused on AI Training

A remote contractor role for a Physical Sciences expert to help train next generation AI systems with real-world scientific input. You will design advanced evaluation tasks, source authentic research artifacts, and craft realistic problem scenarios that reflect complex technical challenges. You’ll document protocols and grading rubrics to ensure reproducibility and high standards of methodological integrity. This position emphasizes domain knowledge over AI experience, focusing on physics, chemistry, or materials science expertise and rigorous quantitative analysis.

30–50/hr
· 10 openings
Methodological RigorRubric FidelityExperimental Reasoning+3 more
Mercor

Mercor

4h agoRemotehourly

Remote Survey Research Expert for AI Evaluation

Mercor is connecting you with a leading AI research organization that needs seasoned survey professionals to evaluate its surveys for methodological quality. You will examine question wording, response scales, and sampling approaches, and articulate why certain designs yield unreliable or biased data. This remote, hourly role suits practitioners from market research, consumer insights, behavioral science, and related fields who have designed and fielded surveys in real projects. Your critical eye will help improve how AI systems think about research design and evidence quality. Candidates with hands-on experience at research agencies or internal insights teams will feel at home.

120–120/hr
· 10 openings
Survey ResearchMarket ResearchConsumer Insights+17 more
Micro1

Micro1

13d agoRemotecontract
Hot

Evaluation Specialist for AI Training and Research

Remote contractor role for Evaluation Specialists or Recent Grads who will help train next‑generation AI systems. You will craft original QA pairs, perform rigorous source triangulation, and create multi‑step questions that require synthesis. The project emphasizes high-quality, real‑world input to influence how models learn, reason, and perform. Prior AI experience isn’t required; deep domain knowledge matters and it’s welcomed.

20–60/hr
· 5 openings
Research & Source TriangulationAttention To DetailWritten Precision+2 more
Micro1

Micro1

1mo agoRemotecontract
Hot

Data Science Domain Expert for AI Evaluation and Prompting

Remote contractor role focusing on AI data science with a domain expert lens. You’ll assess AI outputs, refine prompts, and annotate data to support high-quality model training. The work centers on applying deep domain knowledge to review research-style documents, technical reports, and experiment notes, shaping how models learn and reason. Strong writing, precise attention to detail, and independent research are essential as you contribute to rubric-based evaluations and content quality.

100–200/hr
· 15 openings
Critical ThinkingAnalytical ReasoningAttention To Detail+22 more
Micro1

Micro1

21d agoRemotecontract
Hot

Technical Writer for AI Benchmark Tasks

A remote contractor role focused on shaping AI benchmark work through precise, real-world technical input. You’ll craft expert evaluation tasks using realistic data formats like CSVs, PDFs, and spreadsheets, and assemble source materials such as product specs and code samples. You’ll define clear problem criteria and build detailed rubrics to judge AI outputs for accuracy and audience relevance. Prior experience in regulated or technical fields is valued, but direct AI domain experience isn’t required. Bolded technologies: CSV, PDF, spreadsheets, API references, regulatory documentation.

30–60/hr
· 10 openings
Precision WritingSource SynthesisAudience Calibration+1 more