You will review health AI model outputs covering diagnoses, clinical reasoning, care recommendations, and clinical workflows. Day-to-day work includes scoring responses against expert judgment, identifying unsafe or clinically weak conclusions, and documenting why an output meets or misses medical standards.
The work requires more than checking whether a final diagnosis is plausible. Reviewers must examine the reasoning path, recognize omitted red flags or contraindications, assess the clarity and rigor of recommendations, and flag failure modes that could compromise patient safety. Feedback is structured and written for AI researchers who will use it to improve model performance.
Mercor screens candidates through a 25-minute conversational interview focused on clinical background, current practice, experience, and motivation. The process also includes verification of medical degree, license status, residency completion, and current practice setting, followed by final roster confirmation with the project team.
The project is scheduled to start immediately and run for six weeks, with a part-time commitment of at least 20 hours per week. Collaboration with researchers is asynchronous; the listing requires candidates to be located in the US but does not explicitly state whether the work is fully remote.
Pay band: $150/hr
