What the work involves
You work on Japanese-language data for model training and evaluation. Typical tasks include ranking two or more model responses to a Japanese prompt, rewriting outputs that are grammatical but unnatural, authoring reference translations between Japanese and English, and writing prompts designed to expose weaknesses — keigo and humble/honorable forms, counters, ambiguous subject dropping, on'yomi/kun'yomi errors in rare readings, katakana loanword misuse, and dialect or regional register. Some batches are domain-flavored: business correspondence, legal or medical text, gaming and manga localization, or customer support transcripts. Most work is written; some projects include audio transcription or pronunciation judgments.
The hard part is rarely spotting broken Japanese. It is judging output that reads smoothly to an intermediate learner but signals the wrong social distance, translates a English idiom too literally, or invents a plausible-looking compound that no one writes. Written rationales matter as much as scores — reviewers use them to calibrate the rubric, and thin justifications are the most common reason contributors get cycled off a project.
What the platform screens for
- Native or near-native Japanese, with the ability to articulate why something is off, not only that it is
- Strong English, since rubrics, guidelines, and rationales are usually in English
- Consistency under a rubric — following someone else's definition of "helpful" rather than your own taste
- Depth that holds up under follow-up questioning; the AI interview will push past a first answer
Relevant credentials help: translation or localization experience, JLPT N1 for non-native speakers, teaching, editorial, or subtitling work. Linguistics training is a plus, not a requirement.
Logistics
Fully remote and asynchronous, with work claimed from a queue rather than assigned in shifts. Contributors commonly report 10–20 hours a week when projects are active, with real gaps between batches — treat this as supplemental rather than primary income. Observed rates for this role sit in the $65–98/hr range depending on project, specialization, and calibration performance; pay bands are as observed and not guaranteed. Onboarding usually includes a paid or unpaid calibration set before live work begins.
