AfterQuery
AfterQueryVerified listing
RemoteHot

Graduate Student AI Evaluator - Remote Contract

50–150/hr
Remote
Posted August 29, 2026
Contract
Not Sure? Upload your resume to see every role you match

Listing checked September 13, 2026 · pay as published by AfterQuery

Overview

AfterQuery seeks current master's, PhD, MD, and MD/PhD students to join a remote AI evaluation research project across academic disciplines. You will use the vocabulary and visual conventions of your field, such as spectra, schematics, maps, or clinical images, to judge written and multimodal material. Projects run on a remote, asynchronous schedule, and you decide which windows to accept, with each window asking for about 10 to 20 hours per week. Pay ranges from $50 to $150 per hour, and a project week can bring in $800 to $1,500 based on discipline and seniority.

What You'll Do6

  • 1Evaluate academic content and AI-generated outputs using your subject-matter knowledge.
  • 2Write clear, well-reasoned rationales for each judgment, pointing to evidence in the material.
  • 3Flag errors, ambiguities, and missing context that only a specialist would notice.
  • 4Read and interpret visual or diagrammatic items tied to your discipline, including schematics, spectra, maps, clinical images, artworks, or data charts.
  • 5Validate annotations and evaluations produced for your domain.
  • 6Complete a defined set of annotation tasks inside each project window, with 10 to 20 hours per week.

Requirements10

  • 1Enrolled in a PhD, master's, MD, or MD/PhD program at the time of application.
  • 2Strong English writing skills and the ability to explain domain-specific reasoning in clear prose.
  • 3Deep command of the visual conventions and technical terms used in your field, such as reading spectra, interpreting schematics, analyzing maps, evaluating clinical images, or formal visual analysis.
  • 4Reliable internet access and the ability to work on remote, asynchronous tasks.
  • 5Candidacy status or beyond (qualifying exams passed, or MS3/MS4 clinical rotations in progress).
  • 6Research experience with image-based or visual data, including microscopy, imaging, diagrammatic modeling, CAD, GIS, spectroscopy, or archival visual materials.
  • 7Prior work in annotation, data labeling, or AI evaluation.
  • 8Affiliation with a top research university or program in your discipline.
  • 9Comfort with AI tools, large language models, or multimodal systems.
  • 10Training in quantitative methods across science, social science, business, or engineering.

Who Should Apply

The ideal candidate is a current graduate or medical student who can read the visual language of a discipline, explain a judgment in plain English, and spot errors a general reviewer would miss. You may be in a PhD, master's, MD, or MD/PhD program, and you likely have research that involves images, diagrams, spectra, maps, or clinical material. This role suits people who want flexible, project-based work that uses their academic training rather than unrelated gig tasks. Applicants who lack strong English writing samples or cannot describe their visual reasoning often score low, as do those who treat the work as generic data entry. The role is less suitable for candidates who need fixed weekly hours, close supervision, or tasks outside their domain.

Salary Insight

The listing advertises $50.00 to $150.00 per hour. A full project week can pay $800 to $1,500 depending on discipline and seniority, and top candidates may qualify for premium rates. Rates at the upper end track with advanced graduate standing, specialized visual-data experience, or prior AI evaluation work.

Location

Typeremote
LocationRemote
This is a remote position

Compensation

$50–150/hr

Required Skills

ai evaluationdata annotationdata labelingmultimodal ailarge language modelsacademic researchtechnical writingvisual analysisimage interpretationspectroscopymicroscopygiscadclinical imagingschematic interpretationmap analysisquantitative methodsenglish writing

Application Tip

In your application, name the exact visual or technical material you can evaluate (for example, spectra, schematics, maps, clinical images, or archival works) and pair it with a short writing sample that shows your reasoning. Quantify your research or annotation experience where possible, and note your program stage, such as candidacy, MS3/MS4, or qualifying exams completed, because those details map to the preferred qualifications and premium rates.

Share:

See NearSkill jobs more often in your search

Application & verification flow

  1. 1Instant rubric match

    Your resume is scanned against this role’s requirements to check qualification fit.

  2. 2Screened before the employer sees it

    Only profiles that clear screening are passed on.

  3. 3Outcome by email

    We notify you at the address on your resume once the screening is reviewed.

Test your fit score before applying

Similar open positions

Explore active roles that match your skills and interests.

AfterQuery

AfterQuery

24d agoRemoteContract
Hot

Undergraduate AI Evaluation Researcher - Remote

AfterQuery hires undergraduates from any discipline for a remote, contract AI evaluation project. You will assess, write, and validate content that needs real academic knowledge, from spectra and schematics to clinical images and maps. The listing notes a 40-hour weekly commitment, while individual project windows ask for about 10-20 hours per week. Work runs on an async schedule, and you choose which projects to accept. Pay ranges from $50 to $150 per hour, and project-week earnings can reach $800-$1,500 based on discipline and seniority.

50–150/hr
AI EvaluationData AnnotationData Labeling+14 more
Handshake AI

Handshake AI

4d agoRemotePart-Time

AI Evaluation Specialist - Remote Fellowship

Handshake AI hires students and graduates to support AI research with leading labs. You will use your academic background to help Large Language Models (LLMs) perform better in specific subjects. Projects run in a remote, async format, and the work includes writing domain-specific prompts and judging LLM responses. No minimum hour commitment applies, and placement depends on project availability. The fellowship is part-time and pays $17 to $30 per hour.

17–30/hr
Large Language ModelsLLM EvaluationPrompt Development+8 more
Handshake AI

Handshake AI

4d agoRemotePart-Time

Software Engineer - AI Code Evaluation (Remote)

Handshake AI brings experienced software engineers into part-time contract projects that judge AI-generated code and technical explanations. Contributors review programming outputs, score them against engineering standards, and write structured critiques that help models handle system design and coding tasks better. The program runs year-round and places people on projects as client needs arise, so you can keep your main job while taking assignments. Most contributors log about 5 to 20 hours per week during active projects, and the schedule stays flexible. Pay reaches $65.00 per hour.

Up to 65/hr
PythonJavaC+++15 more
AfterQuery

AfterQuery

24d agoRemoteContract
Hot

PhD Domain Expert - Remote AI Research Contract

AfterQuery seeks PhD-level researchers for Project Iter, a remote contract engagement that runs on a project-by-project basis. You bring doctoral training in a quantitative, scientific, technical, or humanities field and apply it to applied research tasks, from literature reviews to experiment design. Assignments span 2 to 3 weeks and call for about 10 hours per week, with some projects requiring 10 to 20 hours. Work is asynchronous, so you choose when to contribute while supporting model evaluation, data curation, and benchmarking. Pay ranges from $120 to $180 per hour, and the role offers direct exposure to an early-stage Y Combinator-backed company.

120–180/hr
Literature ReviewTechnical WritingExperiment Design+13 more
Micro1

Micro1

30d agoRemotecontract
Hot

Evaluation Specialist for AI Training and Research

Remote contractor role for Evaluation Specialists or Recent Grads who will help train next‑generation AI systems. You will craft original QA pairs, perform rigorous source triangulation, and create multi‑step questions that require synthesis. The project emphasizes high-quality, real‑world input to influence how models learn, reason, and perform. Prior AI experience isn’t required; deep domain knowledge matters and it’s welcomed.

20–60/hr
· 5 openings
Research & Source TriangulationAttention To DetailWritten Precision+2 more
AfterQuery

AfterQuery

24d agoRemoteContract
Hot

History Expert - Remote Contract AI Evaluation

AfterQuery needs historians to build the scenarios that frontier AI labs use when they test model judgment. You will work from your own period or region, writing tasks around source analysis and causal reasoning and checking how models handle evidence. The contract runs remote and async, with at least 10 hours each week and a rate of $50 to $100 per hour. Your specialty stays central: rigor matters more than volume, and your evaluations shape how future AI systems interpret the past.

50–100/hr
HistoryHistoriographyPrimary Sources+11 more