Mercor
MercorVerified listing

Clinical Medicine Domain Expert | $70-$110/hr Remote

70–110/hr
Hybrid · Bay Area, CA
Posted August 20, 2026
full-time
10 openings

Overview

A leading AI lab's GenAI team needs a practicing clinician to shape how frontier models reason about real clinical work. You will audit medical knowledge tasks for clinical soundness, write the instruction specs and golden solutions that set the standard for correctness, and design benchmarks that track model improvement. This full-time W-2 role runs through Cincinnatus LLC, with placement inside the client's own tools and workflows. Expect a hybrid schedule in the Bay Area, with on-site days each week, so local residency or a self-funded relocation is required.

What You'll Do5

  • 1Audit clinical knowledge tasks and model outputs for missing behaviors, thin reasoning, unsafe recommendations, and answers that would not survive clinical scrutiny.
  • 2Write instruction specifications and golden solutions that define correct responses to complex clinical problems.
  • 3Design new clinical tasks and evaluation sets that reflect how medicine is actually practiced, including building medicine-specific skills and tools with the research team.
  • 4Collaborate with client researchers and specialists in adjacent fields to keep evaluation standards consistent and translate tacit clinical judgment into explicit criteria.
  • 5Deliver precise written feedback that helps non-clinician team members understand why a model answer is clinically weak or dangerous.

Requirements8

  • 1MD or DO from an accredited medical school and a completed residency in a recognized specialty.
  • 2At least four years of post-residency clinical practice in a defined specialty; residency and fellowship training do not count.
  • 3Active, unrestricted medical license in at least one U.S. state and board certification in your specialty.
  • 4Genuine specialization in a clinical field such as internal medicine, oncology, radiology, emergency medicine, surgery, psychiatry, or a medical subspecialty.
  • 5Progression to a senior level like Attending Physician, Medical Director, Division Chief, Associate or full Professor, or Chief Medical Officer, with real ownership of clinical decisions.
  • 6Hands-on professional experience using large language models and the ability to distinguish well-reasoned answers from plausible-sounding wrong ones.
  • 7Available to commit 40 hours per week for an initial six-month engagement.
  • 8Based in the Bay Area or willing to relocate there at your own cost; able to work on-site multiple days per week.

Who Should Apply

The ideal candidate is a board-certified physician with at least four years of independent post-residency practice and a genuine specialty that shows up in daily decisions. This role is not for residents, fellows, or generalists who lack a strong clinical identity, and it will not work for physicians who cannot come on-site in the Bay Area multiple days per week. Candidates who cannot commit to a 40-hour workweek for six months or who have only used LLMs recreationally will not move forward. Rejection often follows when applicants list '4+ years experience' but cannot prove it with board certification and a current unrestricted license, or when they describe AI experience without concrete examples of evaluating model outputs.

Salary Insight

The rate is $70-$110 per hour, which lands at the senior end for clinical advisory work because it demands board certification, four plus years of post-residency practice, and real AI fluency. The final offer depends on your specialty and the depth of your LLM evaluation experience.

Location

Typehybrid
LocationBay Area, CA
Eligible countriesUnited States

Required Skills

clinical medicineinternal medicineoncologyradiologyemergency medicinesurgerypsychiatryutilization managementclinical informaticsmedical affairslarge language modelsai evaluationbenchmark designgenai

Application Tip

Submit two or three written critiques of AI-generated clinical answers, showing exactly why each one would fail in a real patient encounter. Name your board certification and your years of independent practice at the top of your resume, and list any experience building evaluation sets or instruction guidelines.

Share:

See NearSkill jobs more often in your search

How your application is processed

  1. 1Application received

    Your resume and details are logged the moment you apply.

  2. 2ATS + eligibility screening

    We check your profile against the role’s skills, seniority, and requirements.

  3. 3Employer sees qualified profiles only

    Only candidates who clear screening move forward.

See your fit score for every role

Similar open positions

Explore active roles that match your skills and interests.

Turing

Turing

16d agoRemotecontract

Resident Medical Specialist (MD/DO)

This remote role places licensed physicians at the center of AI evaluation for clinical reasoning. You will design test scenarios, assessment frameworks, and gap analyses for medical AI systems, collaborating directly with AI researchers. The engagement is flexible and can fit around your clinical schedule, requiring up to 30 hours per week over a one-month contract, with possible extensions based on performance. Your hands-on medical judgment will shape how AI models handle real-world diagnostic and management problems.

Competitive salary
Clinical ReasoningEvidence Based MedicineCritical Thinking+9 more
Turing

Turing

15d agoRemotecontract

Medicine Physician (MD/DO/Doctoral study/PhD)

Turing, a San Francisco-based research accelerator, works with frontier AI labs and global enterprises to advance AI systems. This role puts licensed physicians at the center of evaluating how AI handles clinical reasoning. You will design evaluation methods for real medical problems, drawing on evidence-based medicine, and build assessment approaches that capture the nuance of clinical practice. The engagement is remote and flexible, up to 30 hours per week for one month, with possible extension based on performance.

Competitive salary
Clinical ReasoningEvidence Based MedicineClinical Decision Making+4 more
SME Careers

SME Careers

6h agoRemotecontract

Medical Doctor Specialist for AI Content Review (Remote)

Work as a Medical Doctor Specialist SME, remotely on an hourly contract to review AI-generated clinical responses and, when needed, author expert medical content. You will assess clinical reasoning, verify step-by-step problem solving, and ensure responses align with the prompt. Your input helps improve AI models by validating accuracy, safety, and adherence to evidence-based guidelines. Specialists across fields including Internal Medicine, Emergency Medicine, and Cardiology are welcome to join the expert network.

Up to 150/hr
Board CertificationBoard EligibilitySpecialty Training+23 more
Micro1

Micro1

13d agoRemotecontract
Hot

Medical Evaluation Specialist - Remote Clinical QA for AI Training

Remote contractor role for medical professionals who will craft and verify high‑level medical QA pairs to train cutting‑edge AI. You’ll source answers from primary literature and guidelines, document rationale with citations, and design questions that require genuine clinical reasoning. You can expect to refine inputs based on reviewer feedback and evolving project standards. This work leverages your domain knowledge to improve model learning and evaluation, with tasks described and compensated per deliverable. Medical and clinical guidelines are central to every deliverable.

40–90/hr
· 5 openings
Research & Source TriangulationAttention To DetailWritten Precision+3 more
Mercor

Mercor

15d agoBay Area, CAfull-time

Pharmaceutical R&D Domain Expert

A leading AI lab's GenAI team needs a senior pharmaceutical R&D specialist to shape how frontier models reason about drug discovery and development. You will review the quality of pharmaceutical knowledge tasks, author instruction specs and golden solutions, and design benchmarks that track model improvement. A full-time W-2 role via Cincinnatus LLC places you with the client's research and program management teams in the Bay Area. The work is hybrid: you must live in the Bay Area and be on-site several days a week when asked. Remote-only work is not an option, and relocation is not covered.

75–115/hr
· 10 openings
Large Language ModelsGenerative AIPharmaceutical Research+12 more
SME Careers

SME Careers

6h agoRemotecontract

Medical AI Content Expert for Remote Contract Work

This remote, hourly contractor role focuses on validating AI-generated medical responses and producing expert healthcare content. You’ll examine the reasoning steps, explain how evidence supports the conclusions, and deliver precise written feedback. You’ll flag unsafe assumptions, missing contraindications, or misinterpretations of tests to improve model reliability. Grounded in epidemiology and clinical medicine, your evaluations shape the accuracy and clarity of medical data used by leading AI initiatives.

Up to 90/hr
MedicalAI Training Data EvaluationClinical Reasoning+25 more