
Math Expert AI Evaluation Fellowship Remote
Listing checked September 18, 2026 · pay as published by Handshake AI
Overview
Handshake AI recruits Math PhDs to sharpen how AI systems handle mathematical reasoning, proof construction, and technical problem-solving. The work centers on crafting difficult domain questions and checking AI-generated answers for accuracy, logical consistency, and mathematical rigor. Contributors join a remote, project-based fellowship with a schedule they set themselves. Most participants log 5 to 20 hours per week while a project is active, and no prior AI or technical experience is needed. Placement follows current project needs, with chances to join later projects as they open.
What You'll Do8
- 1Design rigorous math questions and prompts that test an AI model's reasoning.
- 2Review AI-generated responses for mathematical accuracy and step-by-step logic.
- 3Flag errors in proofs, calculations, and formal arguments.
- 4Write clear feedback that explains where a response fails or succeeds.
- 5Assess whether a model's solution meets the standards of advanced mathematical proof.
- 6Align questions with advanced math domains such as algebra, analysis, topology, or logic.
- 7Work on a self-directed schedule across active projects.
- 8Apply the same evaluation standards across many model responses.
Requirements8
- 1Enrolled in a PhD program or hold a doctorate in mathematics or a related field.
- 2Postdoctoral researchers qualify.
- 3Strong command of advanced mathematical reasoning and proof-writing.
- 4Background in formal logic.
- 5Ability to judge AI-generated content for accuracy and logical consistency.
- 6No prior AI or technical experience required.
- 7Comfort with remote, project-based contract work and flexible hours.
- 8F-1 students may qualify through CPT or OPT; STEM OPT is not supported.
Who Should Apply
The ideal candidate is a math PhD student, postdoc, or doctorate holder who enjoys writing proofs and finding subtle gaps in logical arguments. You can shape a 5 to 20 hour week around research, teaching, or industry work because the schedule is asynchronous and project-based. People who need guaranteed full-time hours or a permanent academic appointment will not find that here. A common reason applicants score low is weak evidence of advanced proof-writing or formal logic training, even when they have strong computational math skills. Work authorization also matters: F-1 students may qualify through CPT or OPT, but STEM OPT is not supported, and a school CPT course requirement may not match the program.
Salary Insight
Compensation reaches up to $75.00 per hour. The listing does not name a lower bound, so pay may depend on the project and your qualifications. A rate near the top of that range reflects the advanced math expertise this fellowship requires.
Pay and demand for Machine Learning & AI roles
AggregatedTypical pay
$75/hour
This role
up to $75/hr
Most Machine Learning & AI roles pay $55–$100 per hour. This role's pay falls inside that range.
Based on 513 similar roles that publish pay · 90 publish only a top rate; those count at the rate they gave
Rates shown per hour. Yearly and monthly pay converted; one-time fees and non-USD pay are not included.
- Live similar roles
- 570
- Listed in last 30 days
- 252
- Remote
- 97%
Hiring most right now: micro1 (262) · Mercor (79) · SME Careers (60)
Most requested skills · share of roles
- python17%
- technical writing9%
- llm evaluation8%
- data annotation7%
Figures from Machine Learning & AI roles live on NearSkill when this page loaded. A role can close before you apply, so check the listing itself.
Compare your resume against these rolesLocation
Compensation
Up to $75/hr
Required Skills
Application Tip
Lead with your proof-writing and formal logic background. Name two or three advanced math areas where you can design original questions, and describe a time you caught a subtle flaw in a proof. Add your PhD status and work authorization details, such as CPT or OPT eligibility, because placement depends on project needs and supported authorization.
See NearSkill jobs more often in your search
Application & verification flow
1Instant rubric match
Your resume is scanned against this role’s requirements to check qualification fit.
2Screened before the employer sees it
Only profiles that clear screening are passed on.
3Outcome by email
We notify you at the address on your resume once the screening is reviewed.
Similar open positions
Explore active roles that match your skills and interests.

Handshake AI
VerifiedMathematics Expert - AI Evaluation Project
Handshake AI seeks mathematicians and research scientists for part-time contract work that supports AI research. You will craft expert-level math problems drawn from real proof, formalization, and computational workflows. Then you assess model answers for technical accuracy and rigor, and send structured written feedback that helps the system reason more like a working mathematician. The fellowship runs remote and asynchronous, so you choose your hours and carry the work alongside another role.

Handshake AI
VerifiedMathematics Expert (India) Part-Time Remote
Handshake AI offers this fellowship as an ongoing, part-time contract for mathematicians and research scientists who can improve model reasoning in pure and applied math. You judge AI answers for proof rigor, formalization, and computational correctness. The schedule stays open, so you can keep the project alongside a current academic or industry role. Contributions center on real mathematics workflows, from symbolic computation to theorem proving.

Micro1
VerifiedMath Expert (PhD)
This remote contract role puts your PhD-level mathematics expertise to work shaping the next generation of AI systems. You'll craft and refine detailed solutions to complex math problems, helping train models to reason more accurately. No prior AI experience is needed—your deep knowledge of advanced mathematics is what counts, and you'll collaborate with a distributed team of experts.

Micro1
VerifiedMath Expert PhD for AI Training: Remote Contract
A remote contractor role for PhD-level mathematicians to support an AI training project. You’ll craft detailed, accurate responses to undergraduate and graduate level math prompts and create high-quality “golden” answers for model training. Collaboration with a distributed expert team ensures clarity and educational value, while you interpret complex problems into clear explanations. Your domain knowledge in mathematics is the key asset, and no AI background is required.

Handshake AI
VerifiedSTEM Professor AI Fellowship - Remote Part-Time
Handshake AI seeks faculty in mathematics, chemistry, biology, physics, or engineering to join its AI research community as fellows. Fellows work from any location and set their own schedule, committing around 10 to 20 hours each week. The project centers on building domain prompts and judging how well large language models answer in your field. Assignments open throughout the year, though availability depends on the discipline. Project placement follows demand, so start dates vary.

Handshake AI
VerifiedAI Evaluation Specialist - Remote Fellowship
Handshake AI hires students and graduates to support AI research with leading labs. You will use your academic background to help Large Language Models (LLMs) perform better in specific subjects. Projects run in a remote, async format, and the work includes writing domain-specific prompts and judging LLM responses. No minimum hour commitment applies, and placement depends on project availability. The fellowship is part-time and pays $17 to $30 per hour.


