
Applied Health & Medicine Benchmark Specialist | $94-$119/hr Remote
Overview
Medical and health science experts can earn $94-$119 per hour writing or checking difficult multiple-choice problems for an AI research project. Mercor assigns each contractor one of two roles: authoring original exam questions in their specialty, or verifying pre-written items for precision and soundness. Content spans clinical medicine, medical imaging, pharmacovigilance, healthcare management, and rehabilitation. Contributors rate difficulty, produce a step-by-step chain-of-thought solution in markdown, and attach 1-5 citations from journals or guidelines. The work is remote and asynchronous, with a 10+ hour weekly commitment.
What You'll Do7
- 1Write original multiple-choice questions in your medical specialty that require conceptual reasoning rather than memorization.
- 2Ensure every question is self-contained and unambiguous, so all needed information appears inside the problem statement.
- 3Classify each item as Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above).
- 4Develop one correct answer plus nine distractors that are plausible but incorrect at an expert level.
- 5Create a step-by-step chain-of-thought explanation in markdown that outlines the reasoning in a clear, concise way.
- 6Attach 1 to 5 references from peer-reviewed journals or clinical guidelines for every question.
- 7On verification assignments, identify flaws in clarity, completeness, precision, or solvability and document the edits you make.
Requirements5
- 1Hold an MD, DO, or PhD, or be a doctoral candidate in medicine, biomedical science, public health, or a closely related discipline.
- 2A master's degree may qualify if you show exceptional depth in a narrow medical subdomain.
- 3Demonstrate graduate-level command of medical knowledge, clinical reasoning, and biomedical research methods.
- 4Board certification, direct patient-care experience, or peer-reviewed publications in health fields carry extra weight.
- 5Write clear, precise English that can explain complex clinical concepts to expert readers.
Who Should Apply
The ideal candidate has an advanced clinical or research degree and can write or critique multiple-choice questions at a graduate level. This role suits medical professionals who want part-time, schedule-flexible work that supports AI evaluation. Generalist writers or those without a strong background in one of the listed health domains will find the expectations steep. Common rejection reasons include sample questions that are too easy, ambiguous, or lack citations, or an inability to produce nine plausible distractors. Reviewers also score candidates low when their edits are vague or they fail to justify changes to the original item.
Salary Insight
Pay is $94 to $119 per hour, billed hourly. That rate sits at the high end for medical content work and matches the expectation of MD- or PhD-level expertise.
Location
Required Skills
Application Tip
Prepare a sample multiple-choice question with a step-by-step chain-of-thought solution and two peer-reviewed citations. Mercor uses practical submissions to judge whether you can write and verify items at the required rigor.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedApplied Biology Benchmark Specialist
A remote, asynchronous contract for PhD-level biologists to write and audit multiple-choice assessment content for an AI benchmarking initiative. Work divides into two tracks: authoring original questions or verifying existing ones. Topics span pharmaceutical manufacturing, synthetic biology, drug discovery, and agricultural plus environmental biology. Each item includes one correct answer, nine plausible distractors, a Chain-of-Thought solution in markdown, and up to five academic references. The rate sits at $60 to $75 per hour for a minimum of 10 weekly hours.

Mercor
VerifiedApplied Engineering Benchmark Specialist
Expert engineers write and verify advanced multiple-choice assessment content for an AI research initiative. The work splits into two tracks: authoring original questions or validating pre-written ones across domains like VLSI, control science, mechatronics, and biotechnology. Each question demands deep conceptual reasoning, a difficulty rating from medium to expert, and a full chain-of-thought solution. The role is asynchronous, fully remote, and expects at least 10 hours per week.

Micro1
VerifiedMedical Evaluation Specialist - Remote Clinical QA for AI Training
Remote contractor role for medical professionals who will craft and verify high‑level medical QA pairs to train cutting‑edge AI. You’ll source answers from primary literature and guidelines, document rationale with citations, and design questions that require genuine clinical reasoning. You can expect to refine inputs based on reviewer feedback and evolving project standards. This work leverages your domain knowledge to improve model learning and evaluation, with tasks described and compensated per deliverable. Medical and clinical guidelines are central to every deliverable.

Mercor
VerifiedApplied Mathematics Benchmark Specialist
This role supports an AI research initiative by creating and validating advanced mathematics assessments. You choose between two tracks: question authoring (writing original multiple-choice problems) or question verification (checking and editing existing problems). The subject matter spans signal processing, financial mathematics, mathematical economics, combinatorial optimization, and climate modeling. The work is asynchronous and 100% remote, with a commitment of 10 or more hours per week. Each question must meet exacting standards for clarity, rigor, and difficulty calibration.

SME Careers
VerifiedMedical Doctor Specialist for AI Content Review (Remote)
Work as a Medical Doctor Specialist SME, remotely on an hourly contract to review AI-generated clinical responses and, when needed, author expert medical content. You will assess clinical reasoning, verify step-by-step problem solving, and ensure responses align with the prompt. Your input helps improve AI models by validating accuracy, safety, and adherence to evidence-based guidelines. Specialists across fields including Internal Medicine, Emergency Medicine, and Cardiology are welcome to join the expert network.

Mercor
VerifiedApplied Chemistry Benchmark Specialist
Expert chemists will build and audit multiple-choice assessment items for an AI research initiative. The work splits into two tracks: writing original questions or checking and editing pre-written ones. Domains range from polymer chemistry and energy storage to pharmaceutical and food chemistry. Every item demands a difficulty rating, one correct answer, nine subtle distractors, a markdown chain-of-thought solution, and 1 to 5 academic references. The engagement is remote and asynchronous, with a minimum of 10 hours per week.

