AI jobs for doctors and medical experts
Frontier AI labs pay clinicians, coders and researchers to judge whether a model's medical answer is correct, safe, and defensible against the standards their own field is held to.
- Open roles
- 50
- Median rate
- $85/hr
- Publish an hourly rate
- 100%
listings, updated continuously
from 50 published bands
the sample every median here uses
What each medical specialty pays
Every listing we track, grouped by specialty. Rates are read from the platform's own posting — the median is of published hourly bands, and the range runs from the lowest published floor to the highest published ceiling.
| Specialty | Roles | Median | Published range | Distribution |
|---|---|---|---|---|
| Research, Trials & Biostatistics | 7 | $107n=7 | $25–230 | |
| General Medical & AI Evaluation | 3 | $100n=3 | $75–120 | |
| Clinical Practice & Nursing | 11 | $90n=11 | $30–250 | |
| Pharma, Regulatory & Medical Writing | 8 | $88n=8 | $50–210 | |
| Medical Coding, Billing & Payer Operations | 17 | $82n=17 | $25–150 | |
| Adolescent & Mental Health Safety | 4 | $70n=4 | $45–95 |
Open medical roles
The listings behind the figures above, richest published ceiling first.
What medical experts do in AI
Review model output for clinical safety
Judge whether an AI's clinical answer is correct — and whether it would be safe in front of a patient.
Audit coding and payer decisions
Review AI-generated ICD-10, HCC and prior-authorisation outputs the way you would audit a coder's or a payer's work.
Author benchmark cases and rubrics
Write the expert-level questions and grading rubrics a model is measured against, each with a defensible correct answer.
Build regulatory and trial documents
Produce the study-report sections, safety narratives and synthetic trial data that AI agents are trained and tested on.
Join a clinical talent pool
Apply once, pass a screening interview, and be matched to lab projects as they open rather than to one fixed role.
Who this work is for
Backgrounds in demand
- Medical coders, billers & RCM leaders17roles
- Physicians & hospitalists11roles
- Pharmacovigilance, regulatory & medical writers8roles
- Clinical researchers & biostatisticians7roles
- Child & adolescent mental-health clinicians4roles
Commonly asked for
- Clinical reasoning
- Clinical risk assessment
- Severity staging
- ICD-10 & CPT coding
- HCC / risk adjustment
- Medical necessity review
- Payer criteria (InterQual, MCG)
- Clinical documentation
- Pharmacovigilance
- Regulatory writing
- Trial data interpretation
- Rubric-based evaluation
Drawn from the listings themselves and not counted — the skills field on a listing is free text, and no ranking of it would mean anything.
Platforms hiring medical experts
Figures are calculated over the medical listings each platform currently has open, on the same basis as the rate report.
Also join these expert networks
- PROFILE ENTRYEthos
One profile with Ethos, matched to paid opportunities across fields as they open — not an application to a single listing.
Create profile on Ethos → - PROFILE ENTRYHandshake AI
One profile with Handshake AI, matched to paid opportunities across fields as they open — not an application to a single listing.
Create profile on Handshake AI →
How AI expert work actually works
- 7 min read
What Is AI Evaluation Work? The Complete Primer for Domain Experts
AI labs pay professionals to review and grade model outputs in their field. Here's what the work actually involves, why it exists, and why your expertise is the product.
- 6 min read
How to Pass the AI Interview: What Platform Screens Actually Measure
Mercor and micro1 both screen you with an AI interviewer before any human looks at your file. Knowing what the system scores removes most of the anxiety — and most of the failure modes.
- 6 min read
Maximizing Earnings in AI Evaluation: Rates, Tiers, and Moving Up
The spread between entry annotation and expert review is wide. Here's how the pay tiers actually work and the specific moves that shift you into the higher ones.
- 6 min read
Mercor vs micro1: A Comparison for Candidates
The two biggest expert platforms differ in screening, role mix, and who they suit. Here's the practical breakdown — including where each one frustrates people.
Questions, answered plainly
Mostly judgement work: reviewing AI-generated clinical answers for correctness and safety, auditing coding and payer decisions, and writing the benchmark cases and rubrics a model is scored against. The clinical judgement is what is being bought — the AI side is the subject, not the skill.
The median and each specialty's median and published range are at the top of this page, every one with the number of listings behind it. Rates are read from the platform's own posting; where a platform publishes nothing, we show nothing.
Usually not. What listings ask for is verifiable depth in your own field — a licence, a board certification, a coding credential, years in a specialty. Where familiarity with AI tools appears in the eligibility criteria it is typically listed as preferred rather than required, and each role page separates the two.
Most listings name a specific credential — board certification, an RN or physician licence, a professional coding certification, or a regulatory or biostatistics background. Arrangements vary by listing: much of this is project-based contract work paid by the hour, some listings are restricted to a single country, and a few are full-time. Each role page states what that listing says; we publish no remote percentage, because many listings say nothing about location and silence is not the same as “anywhere”.
Other fields
New medical roles, by email
One message on the days medical listings open. Nothing on the days they do not.