
Medical Expert - Remote AI Evaluation & Review
Listing checked September 16, 2026 · pay as published by Mercor
Overview
Mercor recruits health professionals who can supply the medical judgment that top AI research teams cannot generate on their own. The medicine pool is an open call, not a single opening, so it remains live while projects start on short notice. When a project launches, you may write a case from your practice with the presentation, workup, constraints, and reference answer that sets the standard for grading a model's attempt. You may also review model output for clinical reasoning, checking whether the differential follows a sensible order, whether a guideline applies to the patient, and whether the care plan is safe to carry out. Mercor sets rates per project, and most work is remote and part-time.
What You'll Do7
- 1Write clinical cases from your own practice, including the presentation, workup, constraints, and the reference answer that sets the grading standard for a model's response.
- 2Review model output on clinical reasoning, judging whether the differential appears in a logical order and whether the cited guideline fits the patient in front of you.
- 3Flag recommendations that look correct in a textbook but fail in practice because of comorbidities, access limits, or what a care setting can deliver.
- 4Explain your reasoning in clear written notes so researchers can see the basis for each score or correction.
- 5Evaluate treatment plans for safety and feasibility within the given care setting.
- 6Check whether a model applies a guideline to a patient when the guideline's criteria do not match.
- 7Score model attempts against the reference answer and your clinical standards.
Requirements5
- 1A clinical credential and professional experience in a health profession.
- 2Strong written communication, because most tasks require you to explain your clinical reasoning in text.
- 3Comfort with ambiguous tasks and a willingness to flag instructions that lack clarity.
- 4Ability to assess differential diagnosis, treatment planning, and guideline use across your specialty.
- 5Availability for remote part-time project work that may start on short notice.
Who Should Apply
Clinicians and health professionals who carry clinical judgment in any specialty fit this pool, whether you work in internal medicine, surgery, nursing, pharmacy, dentistry, mental health, laboratory science, public health, or medical writing. The strongest candidates write with precision, can defend a differential or treatment plan in plain language, and spot when a guideline does not match a patient's situation. The role suits people who want flexible, remote evaluation work and can join projects on short notice. A standard full-time clinical post, a guaranteed schedule each week, or direct patient care will not match this work. Candidates often lose fit when they submit a resume without confirming their profession and location, or when they give short written answers in the AI interview that do not show how they reason through a case.
Salary Insight
Mercor lists rates from $60 to $180 per hour. The band is wide because the medicine pool covers many professions, and each project sets pay based on scope and depth. A physician on a complex case-writing project may sit near the top of that range, while other health roles or shorter evaluation tasks may land closer to the bottom. The listing does not guarantee a specific rate; the matching project listing names the exact pay before you accept.
Pay and demand for Medical Research & Public Health roles
AggregatedTypical pay
$88/hour
This role
$60–$180/hr
Most Medical Research & Public Health roles pay $70–$125 per hour. This role's pay falls inside that range.
Based on 61 similar roles that publish pay · 5 publish only a top rate; those count at the rate they gave
Rates shown per hour. Yearly and monthly pay converted; one-time fees and non-USD pay are not included.
- Live similar roles
- 79
- Listed in last 30 days
- 50
- Remote
- 99%
Hiring most right now: micro1 (26) · Mercor (24) · AfterQuery (8)
Most requested skills · share of roles
- ai evaluation14%
- clinical reasoning14%
- technical writing11%
- medical writing9%
Figures from Medical Research & Public Health roles live on NearSkill when this page loaded. A role can close before you apply, so check the listing itself.
Compare your resume against these rolesLocation
Compensation
$60–180/hr
Required Skills
Application Tip
Apply with a resume that names your clinical credential, profession, and practice location, then complete the short AI interview, about 20 minutes total. In the written portions, walk through one real case: the presentation, the differential you considered, the guideline you checked, and the plan you chose. That structure shows the exact reasoning Mercor needs for case writing and model grading.
See NearSkill jobs more often in your search
Application & verification flow
1Instant rubric match
Your resume is scanned against this role’s requirements to check qualification fit.
2Screened before the employer sees it
Only profiles that clear screening are passed on.
3Outcome by email
We notify you at the address on your resume once the screening is reviewed.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedMedical Expert (M.D./D.O.) - Remote Part-Time
Mercor keeps this listing open for physicians who want to shape how frontier AI models handle clinical reasoning. The work fits around your practice: you might write a clinical vignette from a real case, complete with presentation, labs, and imaging, then supply the attending-level answer used as a reference. Other projects ask you to grade model output on diagnostic reasoning, checking whether a differential follows a sensible order, whether the model anchors on a first impression, and whether the proposed workup matches guidelines. You also flag recommendations that read well in a textbook but fail the patient in front of you because of comorbidities, contraindications, or access limits. Projects start on short notice, so Mercor draws from this pool when client demand matches your specialty.

Mercor
VerifiedPhysician (MD/DO) - Remote Clinical AI Evaluation
Mercor needs residency-trained physicians for non-clinical projects that shape how clinical AI systems get measured. You will use your clinical background to build grading criteria, review clinical dialogues, and perform reasoning annotation. The role involves no patient care and no live diagnosis. Mercor runs a shared expert pool, so after onboarding you may join several workstreams at once based on your specialty and availability. Assignments shift as priorities change, and you can move between streams.

AfterQuery
VerifiedMedical MD Expert - Remote Contract Role
AfterQuery, a YC-backed startup, needs practicing or recently graduated physicians to build and grade the clinical material that trains its medical AI. You will write case studies and decision scenarios drawn from real practice, then score how model outputs hold up against clinical guidelines and standard care. The work sits at the meeting point of medical education and digital health, with roughly 20 hours a week on a remote, asynchronous schedule. Cases you design directly shape whether these systems give sound advice when clinicians and patients rely on them.

AfterQuery
VerifiedMedical Expert - AI Clinical Reasoning Contract
AfterQuery brings physicians, nurses, PAs and other clinicians into a remote contract role on its Medicine team. You will write tough clinical prompts and score AI answers for accuracy and sound medical reasoning. The work touches clinical diagnosis, treatment planning, pharmacology and patient safety. Plan for 40 hours each week at $50 to $200 per hour, with an assessment before anyone reviews your application.

Mercor
VerifiedPhysician Talent Network
Mercor is building a network of physicians who want to shape the future of AI in medicine. As a member, you'll work on contract projects for leading AI labs, helping train and evaluate models using your clinical expertise. This is an ongoing talent pool — you apply once, and when a relevant project arises (typically 15-30 hours per week), you'll be considered. Fully remote and $110–$250/hour.

Mercor
VerifiedScience Research Expert - AI Evaluation (Remote)
Mercor builds pools of scientific experts who help frontier AI teams judge model work in domains those teams cannot cover on their own. This listing is a standing application for part-time, remote research work, not a single opening. Accepted experts join a science pool, and Mercor invites matching people to specific projects that name the client, rate, hours, and hiring decision. Projects in life, physical, social sciences, math, and policy research have paid $60 to $120 per hour, with scope and depth setting the exact rate. Your role centers on research design, statistical reasoning, and written judgments that become grading rubrics for AI evaluation.


