What the work involves

Projects drawn from this pool fall into three rough shapes. The first is authoring: writing prompts and reference answers in your language where register, idiom, and cultural context have to be right, not merely grammatical — the kind of answer a native speaker would recognise as written by one of their own. The second is grading: reading model output in your language and judging whether a translation preserves meaning and tone, whether phrasing is idiomatic rather than textbook-correct, whether the register fits the speaker and the setting. The third is audio judgment: listening to synthesised voice, music, or sound design and saying whether it is clean, natural, and usable — and when it isn't, naming the specific artefact, prosody error, or mix problem rather than calling it 'off'.

In all three, the written rationale matters as much as the score. A rating with no defensible reasoning attached is close to worthless to a research team, and most annotation guidelines are written to force that reasoning into the open.

What the screen is actually measuring

The application is a resume, a location confirmation, and roughly a twenty-minute AI-led interview. It is looking for verifiable specifics: which language or languages, at what level, from what — upbringing, formal training, professional practice; what you have translated, voiced, engineered, or taught, and for whom. Expect follow-ups that push past the claim. 'Fluent in Portuguese' invites a question about Brazilian versus European register. 'Audio engineer' invites a question about what you hear in a bad TTS sample. Vague credentials tend to unravel one probe in; concrete ones hold up.

It also probes evaluation temperament — how you handle a rubric that doesn't cover the case in front of you, whether you flag ambiguity or quietly guess, and whether you can separate 'I would have phrased it differently' from 'this is wrong.'

Logistics and honest expectations

  • Remote and asynchronous. Work is done on your own schedule against project deadlines; some audio tasks require a quiet room and decent monitoring headphones.
  • Pay: listings in this field on Mercor have posted at $35–50/hr, set per project by scope, language, and depth. That is an observed range on past postings, not a guarantee.
  • Hours: project-dependent, typically part-time and flexible, often alongside other work.
  • This listing returns no decision. You are added to a pool and contacted when a project needs your language or specialty — possibly within a week, possibly several months. Completing the domain expert interview and joining every network you qualify for improves match odds; you sign up once.