All AI training jobs

Mercor · Finance & specialist

Applied Health & Medicine Benchmark Specialist - review AI outputs in your specialty

Listed on Mercor as “Applied Health & Medicine Benchmark Specialist

$94-$119/hrRemoteContractPaid in USD
ShareWhatsAppTelegramEmail

What this actually is

You bring your specialist expertise to AI evaluation. The shape of the work varies but the pattern is the same: review outputs, rate quality, write prompts, flag errors. The platform title (Applied Health & Medicine Benchmark Specialist) reflects the rate band and the expertise required, not the day-to-day work.

Advertisement

Can you do this on your visa?

F-2 / F-4 / F-5 / F-6: open. E-1 to E-7: needs concurrent-employment permit. D-2 / D-4 students: S-3 permit, 20 hr/week cap. D-10 / D-8: case by case.

Korean tax on USD income

First 5 years in Korea: foreign-source income only taxed if remitted into Korea. After year 5: worldwide income. Full tax guide.

Original posting from Mercor

Role Overview

We are seeking expert medical and health science professionals to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core health and medicine domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types:

  • Question Authoring - Create original, challenging multiple-choice questions in your area of medical expertise, rate their difficulty, and submit them for review.
  • Question Verification - Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made.

Health & Medicine Domains Covered

Clinical Medicine & Surgery, Medical Imaging & Diagnostics, Pharmacovigilance, Healthcare Management & Economics, Rehabilitation and Allied Health.

Key Responsibilities

  • Author original health and medicine questions that test deep conceptual understanding, not surface-level recall
  • Ensure questions are unambiguous, self-contained, and precisely defined - all necessary information must be in the problem statement
  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above)
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers
  • Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format
  • Supply 1-5 academic references per question from reputable sources (peer-reviewed journals, clinical guidelines)
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made

Ideal Qualifications

  • MD, DO, PhD, or doctoral candidate in Medicine, Biomedical Sciences, Public Health, or a closely related field
  • Master's degree considered for candidates with exceptional depth in a specific subdomain
  • Strong command of graduate-level medical knowledge, clinical reasoning, and biomedical research methodology
  • Board certification, clinical experience, or research publications in health fields is a strong plus
  • Excellent written English and ability to express complex ideas clearly and concisely

More About the Opportunity

  • Expected commitment: 10+ hours/week
  • Asynchronous, fully remote work

Quoted from Mercor’s public listing on 2026-09-08. We don’t edit platform copy; honest framing is in the title and the “what this actually is” block above.

Apply on Mercor