What the work involves
This is search-and-verify work, not writing work. On a typical shift you might author a question that requires stitching together three or four independent sources — a filing, a dataset, an archived page, a non-English press release — and then produce the full retrieval path: queries used, pages visited, the exact passage that settles the answer. Other tasks run in the opposite direction: a model has already produced a research answer with citations, and your job is to check every link, flag the citation that supports a weaker claim than the one it's attached to, and decide whether the reasoning chain actually closes or just looks like it does.
The common thread is provable ground truth. Tasks are rejected when the answer is contestable, when it hinges on a paywalled source nobody can re-check, or when a competent searcher can find it in one query. Expect to spend real time on difficulty calibration — the interesting band is questions that take a skilled human twenty minutes and defeat a model outright.
What the platform screens for
- A verifiable research background. Mercor's screen is credential-anchored: a graduate degree, a publication record, a journalism or legal-research or analyst track record. Domain doesn't matter; depth does.
- Demonstrated search fluency. Interviewers probe how you actually find things — operator use, archive and database habits, how you handle a dead end or a source in a language you don't read.
- Source discipline. Can you distinguish primary from secondary, tell when a citation is decorative, and say plainly when the evidence doesn't support a conclusion?
- Written precision under follow-up. The AI interviewer pushes on vague claims; specifics survive, generalities don't.
Logistics
Fully remote and asynchronous — tasks are claimed from a queue with per-task or weekly deadlines rather than fixed hours. Most contributors commit 10–20 hours a week, and availability questions in the screen are taken literally, so answer with the hours you will actually deliver. Onboarding usually includes a calibration batch reviewed against a rubric before volume opens up; pay is hourly against logged, reviewed work.