
AI Evaluation Specialist - Remote Fellowship
Listing checked September 18, 2026 · pay as published by Handshake AI
Overview
Handshake AI hires students and graduates to support AI research with leading labs. You will use your academic background to help Large Language Models (LLMs) perform better in specific subjects. Projects run in a remote, async format, and the work includes writing domain-specific prompts and judging LLM responses. No minimum hour commitment applies, and placement depends on project availability. The fellowship is part-time and pays $17 to $30 per hour.
What You'll Do6
- 1Write domain-specific prompts that test how well a model handles ideas from your academic field.
- 2Evaluate LLM responses for accuracy, reasoning, and how well they follow instructions.
- 3Apply detailed project instructions or rubrics to keep each judgment consistent.
- 4Record your evaluations so researchers can review the reasoning behind them.
- 5Complete short-term tasks on your own schedule with no minimum hours.
- 6Switch between topics, guidelines, and project needs as new AI training work opens.
Requirements6
- 1Current associate, bachelor's, or master's student, or a graduate who already holds a degree.
- 2Strong attention to detail and a record of following detailed instructions with care.
- 3Comfort working alone in an asynchronous, remote setting.
- 4Ability to reason through complex material in at least one academic or technical domain.
- 5F-1 students may qualify through CPT or OPT; STEM OPT does not qualify. Confirm eligibility with your Designated School Official.
- 6If your school requires a CPT course, this program may not meet that requirement.
Who Should Apply
The role suits current associate, bachelor's, or master's students and recent graduates who want flexible, project-based AI work. You need a sharp eye for detail and the patience to follow long instructions without cutting corners. Candidates who treat this as a way to apply real subject knowledge to model evaluation tend to fit best. The role is less suitable for anyone who needs guaranteed weekly hours or a permanent full-time schedule. Two common issues that hurt applicants: failing to show how they handle detailed rubrics, and missing work authorization limits, since STEM OPT is not supported.
Salary Insight
Handshake AI lists $17 to $30 per hour for this fellowship. That range matches entry-level, project-based AI evaluation work, and your exact rate may depend on the project you join and your academic background.
Pay and demand for Machine Learning & AI roles
AggregatedTypical pay
$75/hour
This role
$17–$30/hr
Most Machine Learning & AI roles pay $55–$100 per hour. This role's pay sits below that range.
Based on 509 similar roles that publish pay · 90 publish only a top rate; those count at the rate they gave
Rates shown per hour. Yearly and monthly pay converted; one-time fees and non-USD pay are not included.
- Live similar roles
- 566
- Listed in last 30 days
- 250
- Remote
- 97%
Hiring most right now: micro1 (259) · Mercor (79) · SME Careers (60)
Most requested skills · share of roles
- python17%
- technical writing9%
- llm evaluation8%
- ai evaluation7%
Figures from Machine Learning & AI roles live on NearSkill when this page loaded. A role can close before you apply, so check the listing itself.
Compare your resume against these rolesLocation
Compensation
$17–30/hr
Required Skills
Application Tip
Name your degree, major, or strongest subject area, then give one concrete example of a time you followed a detailed rubric and caught an error others missed. If you are an F-1 student, state whether you have CPT or OPT eligibility up front because work authorization often decides fit.
See NearSkill jobs more often in your search
Application & verification flow
1Instant rubric match
Your resume is scanned against this role’s requirements to check qualification fit.
2Screened before the employer sees it
Only profiles that clear screening are passed on.
3Outcome by email
We notify you at the address on your resume once the screening is reviewed.
Similar open positions
Explore active roles that match your skills and interests.

Handshake AI
VerifiedComputer Science Expert - Remote AI Fellowship
Handshake AI runs a year-round fellowship that brings doctoral students and graduates in Computer Science together to improve Large Language Models. Fellows work from any location and set their own hours, with no minimum weekly commitment. Assignments focus on writing domain-specific prompts and reviewing model outputs for accuracy and reasoning. Project availability shifts by domain, so placement depends on current openings.

Handshake AI
VerifiedEngineering Expert - Remote AI Fellowship
Handshake AI brings together PhD and master's level engineers, with a focus on hardware engineering, to sharpen how large language models handle specialized technical subjects. The AI Fellowship runs all year, and project openings shift by domain and availability. Fellows work on a remote, asynchronous schedule for about 10 to 20 hours each week. Typical project tasks include writing domain-specific prompts and judging LLM responses, while also researching topics that matter to your field with AI tools at your side. The program accepts U.S.-based doctoral students, postdocs, and recent graduates who hold valid work or training authorization.

Handshake AI
VerifiedEngineering Professional - AI Fellowship (Remote)
Handshake AI hires experienced engineers for a part-time fellowship that sharpens how AI models handle professional engineering work. You will write industry-specific prompts and review LLM responses for accuracy, safety, and technical depth. The role is remote and asynchronous, with flexible hours and no minimum weekly commitment. Projects open across mechanical engineering, electrical engineering, and biomedical engineering throughout the year, and placement depends on current project availability.

Handshake AI
VerifiedRemote AI Evaluation Specialist in Canada
Handshake AI hires students and degree holders across Canada to support research programs run with leading AI labs. Contributors use their academic background to sharpen how Large Language Models (LLMs) handle specialized subjects. The work happens on remote, asynchronous AI training projects, and you set your own hours with no minimum commitment. Typical tasks range from writing domain-specific prompts to reviewing and grading LLM responses. Project placement depends on what work is open at the time.

Handshake AI
VerifiedLegal Professionals - Remote AI Research Fellowship
Handshake AI invites legal professionals to support an AI research project. Your legal knowledge helps models read and respond to professional tasks with greater accuracy. The program runs year-round, and project openings appear based on demand. Contributors work remote and asynchronous, set their own hours, and have no minimum commitment. Assignments often involve drafting industry-specific prompts and reviewing LLM responses.

Handshake AI
VerifiedCommunication Professionals - Remote Fellowship
Handshake AI seeks communication professionals for an AI research project. The role centers on industry-specific prompts and LLM response evaluation, using your editorial or journalism background to sharpen how AI handles professional language. Contributors work remote on an asynchronous schedule, with flexible hours and no minimum commitment. The program runs year-round, though project slots appear when new work opens. Placement depends on current project availability.


