
Engineering Expert - Remote AI Evaluation
Listing checked September 16, 2026 · pay as published by Mercor
Overview
Mercor supports research teams behind well-known AI models by supplying engineering judgment that models cannot generate on their own. Your design standards and analysis habits become the rubric used to grade a model's engineering output. The company keeps a standing, part-time listing for engineers who want AI evaluation and training work, not a single opening. Projects can start with short notice, so Mercor draws from this pool when a client needs a discipline such as structural engineering, mechanical engineering, or electrical engineering. Past engineering projects posted at $60.00 to $125.00 per hour.
What You'll Do7
- 1Write design problems from your own engineering work, including the loads, constraints, tolerances, and the reference answer used to grade a model's attempt.
- 2Review model-written analysis and decide whether it picked the governing failure mode, held its assumptions within the claimed regime, and applied the safety factor to the right quantity.
- 3Mark cases where a model's design passes calculation but fails in practice because of manufacturability, installation, corrosion, or what a standard requires.
- 4Explain your reasoning in writing so project leads and clients can follow the basis for each grade.
- 5Flag model outputs that look correct on paper yet miss field conditions or code requirements.
- 6Work within one or more engineering disciplines, such as civil, chemical, materials, aerospace, industrial, robotics, or environmental engineering.
- 7Complete a short AI interview and background verification before joining the engineering pool.
Requirements6
- 1A degree in an engineering discipline.
- 2Professional experience in design, analysis, or both.
- 3Clear written communication, since most tasks ask you to defend a grade or explain an assumption.
- 4Comfort with ambiguous or incomplete task instructions.
- 5Willingness to flag unclear instructions instead of guessing.
- 6A design you worked on that was built, plus knowledge of how it performed after construction or deployment.
Who Should Apply
Engineers with a degree and real design or analysis experience fit this pool best, and those who have watched a design move from drawings to construction or production and learned what changed. Written skill matters because most tasks ask you to justify a grade or explain an assumption in plain technical language. The role suits people who can handle vague prompts without freezing and who will say when instructions lack detail. Candidates who only have academic research experience, or who dislike writing explanations, tend to score lower for this work. A common rejection reason is a resume that lists coursework and tools but shows no design the candidate owned or no outcome from a built system. Another is failing to flag ambiguity, which makes grading unreliable for clients.
Salary Insight
Mercor lists $60.00 to $125.00 per hour for engineering projects. Clients set the rate per project by scope and depth, so complex design or analysis work tends to land near the top of that band while smaller grading tasks sit lower. This part-time work pays through the specific project listing when Mercor invites you.
Pay and demand for Machine Learning & AI roles
AggregatedTypical pay
$75/hour
This role
$60–$125/hr
Most Machine Learning & AI roles pay $55–$100 per hour. This role's pay falls inside that range.
Based on 596 similar roles that publish pay · 98 publish only a top rate; those count at the rate they gave
Rates shown per hour. Yearly and monthly pay converted; one-time fees and non-USD pay are not included.
- Live similar roles
- 671
- Listed in last 30 days
- 318
- Remote
- 96%
Hiring most right now: micro1 (291) · Mercor (99) · Handshake AI (64)
Most requested skills · share of roles
- python16%
- technical writing9%
- llm evaluation7%
- ai evaluation6%
Figures from Machine Learning & AI roles live on NearSkill when this page loaded. A role can close before you apply, so check the listing itself.
Compare your resume against these rolesLocation
Compensation
$60–125/hr
Required Skills
Application Tip
Send a resume that names at least one design you owned, the loads or constraints you worked within, and what happened after it was built or installed. Add a short writing sample or bullet that explains a technical decision, because the AI interview and later grading tasks test how well you justify an answer. Confirm your work location and finish the 20-minute domain expert interview to raise your match odds.
See NearSkill jobs more often in your search
Application & verification flow
1Instant rubric match
Your resume is scanned against this role’s requirements to check qualification fit.
2Screened before the employer sees it
Only profiles that clear screening are passed on.
3Outcome by email
We notify you at the address on your resume once the screening is reviewed.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedMechanical Engineering Expert (Remote, Part-Time)
AI research teams rely on Mercor for mechanical engineering judgment that models cannot generate alone. Your design standards and analysis habits become the rubric that grades frontier AI. Mercor keeps a standing listing for mechanical engineers who want AI training and evaluation work, even when no project is active. When a project starts, matching engineers from the pool receive an invite that names the rate, hours, and client.

Mercor
VerifiedElectrical Engineering Expert - Remote Part-Time
Mercor maintains a standing roster of electrical engineers to support AI research teams that need human judgment on hardware problems. The listing does not tie to one active project. Instead, it lets you join a pool that Mercor draws from when a client needs help writing problems, grading model output, or evaluating design choices. Work can involve circuit design review, systems analysis, noise and stability margins, and checks for whether a model's answer survives real board conditions like parasitics or tolerance stack-up. Rates for past projects ranged from $70 to $120 per hour, with scope and depth setting the final number.

Mercor
VerifiedScience Research Expert - AI Evaluation (Remote)
Mercor builds pools of scientific experts who help frontier AI teams judge model work in domains those teams cannot cover on their own. This listing is a standing application for part-time, remote research work, not a single opening. Accepted experts join a science pool, and Mercor invites matching people to specific projects that name the client, rate, hours, and hiring decision. Projects in life, physical, social sciences, math, and policy research have paid $60 to $120 per hour, with scope and depth setting the exact rate. Your role centers on research design, statistical reasoning, and written judgments that become grading rubrics for AI evaluation.

Mercor
VerifiedSoftware Engineer Expert (Remote) Part-Time
Mercor keeps a standing pool of software engineers who help grade and train frontier AI models. The work is part-time and remote, and projects span frontend, backend, mobile, embedded, DevOps, QA, and security engineering. Assignments range from writing original engineering problems with reference answers to reviewing model-written code for edge cases, complexity, and production failure modes. Past software engineering projects on Mercor paid between $70 and $150 per hour, with scope and depth setting the exact rate.

Mercor
VerifiedGeneralist AI Evaluation Expert (Remote)
Mercor runs a standing pool for people who want part-time work in AI training and model evaluation. The team connects generalists to projects that need careful human judgment, from writing prompts and reference answers to grading how well a model handles general reasoning. Contracts pay $40 to $70 per hour and open on short notice, so the listing stays active even when no project is live. You apply once, complete a short interview, and join a network that clients draw from when a matching task appears.

Mercor
VerifiedMechanical Engineer - AI Lab Project (US Remote)
Mercor pairs seasoned mechanical engineers with a frontier AI lab's GenAI group for a remote, W-2 role through Cincinnatus LLC. You will judge how well large language models handle real engineering problems, from FEA and CFD analyses to drawings, tolerances, and design reviews. The work centers on quality review, instruction specs, worked reference solutions, and benchmark sets that test genuine domain reasoning. Expect a 2 to 3 month initial engagement at 40 hours per week, with a strong chance of extension if the collaboration goes well. Pay runs $60 to $90 per hour.


