What the work actually involves
You write prompts that sit on the line your own field lives on. A method-validation question about limit of detection for an organophosphate is routine professional practice; the same numbers, framed differently, describe the concentration at which a screen stops seeing something. Each prompt you submit is labelled benign, dual-use or adversarial, and the label has to be defensible. You then evaluate what the model returns against a written policy standard and decide whether it handled the request correctly — including the failure mode most people forget, where a model refuses a legitimate GC-MS column-selection question and leaves a working analyst without an answer.
Every item ends with a reference answer: what a correct response looks like, and the technical reasoning for why that boundary falls where it does. This is the writing-heavy half of the job. Rationales are read by policy staff and reviewers who are not chemists, so "any competent analyst would know this is fine" is not a rationale. Prior technical writing, published method papers, or expert witness reports are strong signals, and the application asks for a sample or link.
What the platform screens for
- Hands-on method development and validation, not just instrument operation — whether you can explain LOD versus LOQ versus reporting limit without reaching for a textbook, and what each implies operationally.
- Calibrated judgment on dual-use framing: can you articulate why one question about spectral library coverage is routine and a near-identical one is fishing?
- Writing that a non-specialist can follow. Screeners probe this directly and it is where a large share of technically strong applicants fall out.
- Clean disclosure on obligations. Work here must not draw on classified or export-controlled material, NDA-covered information, or anything under prepublication review. Holding such obligations is not disqualifying — concealing them is; the team scopes work around them.
Logistics
Fully remote and asynchronous, task-based rather than shift-based. Observed pay on this panel has been $65–75 per task; that is what contributors have reported, not a guarantee, and effective hourly rate depends heavily on how fast you write rationales. Volume comes in batches and can be uneven. You will be reading and writing about misuse scenarios in chemistry for sustained stretches; the team briefs experts on this beforehand and you can pause or withdraw from a batch without penalty.