The work

You will adapt computational work you have personally run or debugged into research-grade evaluation tasks. Suitable source material includes simulations, numerical solvers, model fitting, data reduction, analysis pipelines and inverse problems. Each task contains a main problem, connected subproblems, a reference solution and tests that assess the scientific result.

Evaluation design

You will define scientifically meaningful metrics without publishing exact passing thresholds. Tasks should be difficult because of the underlying science, not because instructions, inputs or assumptions are missing. Submissions go through structured review and may require revisions; reviewing and calibrating other researchers' tasks against a short rubric is optional.

What the platform screens for

Eligible researchers are active PhD researchers or research-active Master's students or graduates with the stated research record. Working proficiency in Python is required, including the ability to run and debug your own computational work. The screen is likely to examine one concrete research workflow in depth, how you would make it reproducible and testable, and how you distinguish model failure from ambiguity in the task.

Practical details

This is remote, flexible independent-contractor work with no fixed weekly schedule or cap on task volume. A task typically takes about two hours, although review and revision needs may vary, and accepted experts undergo standard background and identity verification. The observed band is not guaranteed and may depend on task approval and platform terms.

Pay band: $60/hr