
Applied Engineering Benchmark Specialist | $61-$77/hr Remote
Overview
Expert engineers write and verify advanced multiple-choice assessment content for an AI research initiative. The work splits into two tracks: authoring original questions or validating pre-written ones across domains like VLSI, control science, mechatronics, and biotechnology. Each question demands deep conceptual reasoning, a difficulty rating from medium to expert, and a full chain-of-thought solution. The role is asynchronous, fully remote, and expects at least 10 hours per week.
What You'll Do7
- 1Draft original engineering questions that probe conceptual depth rather than memorized facts.
- 2Edit and validate existing questions for clarity, completeness, and technical precision.
- 3Assign each question a difficulty tier: medium, hard, or expert based on undergraduate and graduate levels.
- 4Construct one correct answer plus nine plausible distractors for each multiple-choice item.
- 5Write step-by-step chain-of-thought solutions in markdown, with concise intermediate reasoning.
- 6Attach one to five academic references from reputable journals or university repositories.
- 7Document every change made during verification and justify the reasoning.
Requirements5
- 1PhD or doctoral candidate in engineering or a closely related discipline.
- 2Master's degree holders with exceptional subdomain depth are also considered.
- 3Graduate-level mastery of applied mathematics and domain-specific engineering standards.
- 4Professional engineering (PE) licensure or relevant industry experience strongly preferred.
- 5Clear written English and the ability to express complex ideas with concise phrasing.
Who Should Apply
The ideal candidate has a PhD or deep graduate-level command of one of the listed engineering domains, with a proven ability to reason with rigor under ambiguity. This role suits someone who enjoys meticulous technical writing and can defend their answer choices with precise logic. It is less suitable for engineers who prefer building hardware or software over reviewing conceptual problems. Applications often fall short when candidates lack formal graduate training, when they submit vague question drafts without a clear single answer, or when their references do not come from credible academic sources.
Salary Insight
The contract pays $61.00 to $77.00 per hour, based on the specific task and the engineer's depth of expertise. This rate sits above typical freelance assessment work, reflecting the advanced academic rigor expected from PhD-level domain experts.
Location
Required Skills
Application Tip
Submit a sample engineering question in your domain with a complete chain-of-thought solution and three academic references. A concrete example helps reviewers verify your ability to produce rigorous, self-contained problems.
See NearSkill jobs more often in your search
How your application is processed
1Application received
Your resume and details are logged the moment you apply.
2ATS + eligibility screening
We check your profile against the role’s skills, seniority, and requirements.
3Employer sees qualified profiles only
Only candidates who clear screening move forward.
Similar open positions
Explore active roles that match your skills and interests.

Mercor
VerifiedApplied Mathematics Benchmark Specialist
This role supports an AI research initiative by creating and validating advanced mathematics assessments. You choose between two tracks: question authoring (writing original multiple-choice problems) or question verification (checking and editing existing problems). The subject matter spans signal processing, financial mathematics, mathematical economics, combinatorial optimization, and climate modeling. The work is asynchronous and 100% remote, with a commitment of 10 or more hours per week. Each question must meet exacting standards for clarity, rigor, and difficulty calibration.

Mercor
VerifiedApplied Chemistry Benchmark Specialist
Expert chemists will build and audit multiple-choice assessment items for an AI research initiative. The work splits into two tracks: writing original questions or checking and editing pre-written ones. Domains range from polymer chemistry and energy storage to pharmaceutical and food chemistry. Every item demands a difficulty rating, one correct answer, nine subtle distractors, a markdown chain-of-thought solution, and 1 to 5 academic references. The engagement is remote and asynchronous, with a minimum of 10 hours per week.

Mercor
VerifiedApplied Biology Benchmark Specialist
A remote, asynchronous contract for PhD-level biologists to write and audit multiple-choice assessment content for an AI benchmarking initiative. Work divides into two tracks: authoring original questions or verifying existing ones. Topics span pharmaceutical manufacturing, synthetic biology, drug discovery, and agricultural plus environmental biology. Each item includes one correct answer, nine plausible distractors, a Chain-of-Thought solution in markdown, and up to five academic references. The rate sits at $60 to $75 per hour for a minimum of 10 weekly hours.

Mercor
VerifiedApplied Legal Benchmark Specialist
Legal experts with a JD, LLM, or SJD will craft and review multiple-choice questions for an AI research initiative. The work is remote and asynchronous, with two possible tracks: authoring original questions or verifying pre-written ones. Each question targets a core law domain such as intellectual property, securities, or antitrust. Experts rate difficulty from medium to expert and provide chain-of-thought solutions with references.

Micro1
VerifiedMechanical Engineering Professor for Remote Contract Teaching and Assessment
Senior mechanical engineering educators join a remote contractor project to help train next‑gen AI systems. You’ll create exam‑quality content in areas like solid and structural mechanics, dynamics and vibration, and thermal-fluids, including challenging questions and thorough answer rationales. The work centers on validating AI materials and collaborating asynchronously with other experts to uphold rigorous academic standards. Expect to submit a minimum number of tasks weekly and use digital workflow tools to keep records of assignments.

Mercor
VerifiedApplied Health & Medicine Benchmark Specialist
Medical and health science experts can earn $94-$119 per hour writing or checking difficult multiple-choice problems for an AI research project. Mercor assigns each contractor one of two roles: authoring original exam questions in their specialty, or verifying pre-written items for precision and soundness. Content spans clinical medicine, medical imaging, pharmacovigilance, healthcare management, and rehabilitation. Contributors rate difficulty, produce a step-by-step chain-of-thought solution in markdown, and attach 1-5 citations from journals or guidelines. The work is remote and asynchronous, with a 10+ hour weekly commitment.

