Handshake AI
Handshake AIVerified listing
Remote

Remote AI Evaluation Specialist in Canada

15/hr
Remote
Posted September 22, 2026
Part-Time
Not Sure? Upload your resume to see every role you match

Listing checked September 22, 2026 · pay as published by Handshake AI

Overview

Handshake AI hires students and degree holders across Canada to support research programs run with leading AI labs. Contributors use their academic background to sharpen how Large Language Models (LLMs) handle specialized subjects. The work happens on remote, asynchronous AI training projects, and you set your own hours with no minimum commitment. Typical tasks range from writing domain-specific prompts to reviewing and grading LLM responses. Project placement depends on what work is open at the time.

What You'll Do6

  • 1Build prompts tailored to a chosen subject area so models face challenging questions.
  • 2Assess model answers for accuracy, reasoning, and instruction following.
  • 3Score responses against project guidelines and note where they fail.
  • 4Document findings or feedback for research teams in a clear format.
  • 5Follow detailed task instructions for each short-term project.
  • 6Adapt to new domains and project briefs as assignments change.

Requirements6

  • 1Current enrollment in an Associate's, Bachelor's, or Master's program, or a completed degree.
  • 2Residency in Canada.
  • 3Show exceptional attention to detail and follow detailed instructions with accuracy.
  • 4Comfort with remote, asynchronous work and independent schedules.
  • 5Work with leading AI research labs in a remote setting.
  • 6Use reasoning skills strong enough to challenge advanced AI systems.

Who Should Apply

Students and graduates based in Canada who want flexible, project-based AI work fit this role well. Ideal candidates hold or pursue an Associate's, Bachelor's, or Master's degree and can commit to short assignments without a fixed schedule each week. The work rewards people who read detailed instructions and can judge model outputs with logic and care. Candidates who need guaranteed hours, steady long-term employment, or close supervision may find the setup a poor match. Applications often lose momentum when a candidate cannot show Canadian residency or when their reasoning examples stay vague rather than concrete.

Salary Insight

The listed pay is $15.00 per hour. For part-time, project-based AI evaluation work, that rate sits at the entry-level end of the range, which tracks with a fellowship-style program open to current students and recent graduates. The listing does not mention overtime, bonuses, or pay progression, so treat the hourly figure as the full known compensation at this stage.

Pay and demand for Machine Learning & AI roles

Aggregated

Typical pay

$75/hour

This role

$15/hr

Most Machine Learning & AI roles pay $52–$100 per hour. This role's pay sits below that range.

Based on 513 similar roles that publish pay · 90 publish only a top rate; those count at the rate they gave

Typical rangeMedian payThis role

Rates shown per hour. Yearly and monthly pay converted; one-time fees and non-USD pay are not included.

Live similar roles
572
Listed in last 30 days
253
Remote
97%

Hiring most right now: micro1 (261) · Mercor (79) · SME Careers (60)

Most requested skills · share of roles

  • python
    17%
  • technical writing
    9%
  • llm evaluation
    8%
  • data annotation
    7%

Figures from Machine Learning & AI roles live on NearSkill when this page loaded. A role can close before you apply, so check the listing itself.

Compare your resume against these roles

Location

Typeremote
LocationRemote
This is a remote position

Compensation

$15/hr

Required Skills

large language modelsllm evaluationprompt engineeringai trainingresponse gradingdata annotationreasoning assessmentinstruction followingdomain-specific promptingremote collaborationasynchronous workquality review

Application Tip

In your application, state your Canadian location and your current or completed degree level in the first lines. Add one short example of a prompt you would write in your field and the flaw you would look for in an LLM answer. That concrete reasoning sample speaks louder than a general claim about detail.

Share:

See NearSkill jobs more often in your search

Application & verification flow

  1. 1Instant rubric match

    Your resume is scanned against this role’s requirements to check qualification fit.

  2. 2Screened before the employer sees it

    Only profiles that clear screening are passed on.

  3. 3Outcome by email

    We notify you at the address on your resume once the screening is reviewed.

Test your fit score before applying

Similar open positions

Explore active roles that match your skills and interests.

Handshake AI

Handshake AI

6d agoRemotePart-Time

AI Evaluation Specialist - Remote Fellowship

Handshake AI hires students and graduates to support AI research with leading labs. You will use your academic background to help Large Language Models (LLMs) perform better in specific subjects. Projects run in a remote, async format, and the work includes writing domain-specific prompts and judging LLM responses. No minimum hour commitment applies, and placement depends on project availability. The fellowship is part-time and pays $17 to $30 per hour.

17–30/hr
Large Language ModelsLLM EvaluationPrompt Development+8 more
AfterQuery

AfterQuery

25d agoRemoteContract
Hot

Graduate Student AI Evaluator - Remote Contract

AfterQuery seeks current master's, PhD, MD, and MD/PhD students to join a remote AI evaluation research project across academic disciplines. You will use the vocabulary and visual conventions of your field, such as spectra, schematics, maps, or clinical images, to judge written and multimodal material. Projects run on a remote, asynchronous schedule, and you decide which windows to accept, with each window asking for about 10 to 20 hours per week. Pay ranges from $50 to $150 per hour, and a project week can bring in $800 to $1,500 based on discipline and seniority.

50–150/hr
AI EvaluationData AnnotationData Labeling+15 more
Handshake AI

Handshake AI

5d agoRemotePart-Time

Computer Science Expert - Remote AI Fellowship

Handshake AI runs a year-round fellowship that brings doctoral students and graduates in Computer Science together to improve Large Language Models. Fellows work from any location and set their own hours, with no minimum weekly commitment. Assignments focus on writing domain-specific prompts and reviewing model outputs for accuracy and reasoning. Project availability shifts by domain, so placement depends on current openings.

Up to 75/hr
Computer ScienceLarge Language ModelsLLM Evaluation+11 more
AfterQuery

AfterQuery

26d agoRemoteContract
Hot

Undergraduate AI Evaluation Researcher - Remote

AfterQuery hires undergraduates from any discipline for a remote, contract AI evaluation project. You will assess, write, and validate content that needs real academic knowledge, from spectra and schematics to clinical images and maps. The listing notes a 40-hour weekly commitment, while individual project windows ask for about 10-20 hours per week. Work runs on an async schedule, and you choose which projects to accept. Pay ranges from $50 to $150 per hour, and project-week earnings can reach $800-$1,500 based on discipline and seniority.

50–150/hr
AI EvaluationData AnnotationData Labeling+14 more
Micro1

Micro1

1mo agoRemotecontract
Hot

Evaluation Specialist for AI Training and Research

Remote contractor role for Evaluation Specialists or Recent Grads who will help train next‑generation AI systems. You will craft original QA pairs, perform rigorous source triangulation, and create multi‑step questions that require synthesis. The project emphasizes high-quality, real‑world input to influence how models learn, reason, and perform. Prior AI experience isn’t required; deep domain knowledge matters and it’s welcomed.

20–60/hr
· 5 openings
Research & Source TriangulationAttention To DetailWritten Precision+2 more
Handshake AI

Handshake AI

7d agoRemotePart-Time

Financial Services AI Prompt Evaluator (Remote)

Handshake AI recruits financial services professionals for a remote, part-time fellowship that supports AI research. Projects open across the year as different subject areas become available, so placement depends on current needs. Contributors set their own hours and take on tasks such as writing domain-specific prompts and reviewing output from large language models. The program values real-world experience in roles like credit authorizing, financial risk analysis, and bank teller work.

Up to 100/hr
Financial ServicesCredit AuthorizingCredit Checking+6 more