What the work involves
This is a single, self-contained recording task rather than an ongoing engagement. You receive a short script and a written recording guide, read the script aloud into your phone or computer microphone, and upload the file through the platform's submission form. Most contributors finish in under 15 minutes including setup. The audio is used to build speech datasets that need genuine accent diversity — Arabic-influenced English pronounced by people who actually speak that way, not performed by voice actors.
There is no coaching, no direction session, and no back-and-forth with a producer. The guide will specify things like recording in a quiet room, holding a consistent distance from the mic, avoiding background noise and echo, and reading at a natural conversational pace. Following those instructions exactly is the whole job. Submissions that clip, distort, or have audible traffic, fans, or family conversation behind them are typically rejected and get sent back for a re-record.
What the platform screens for
- Accent authenticity. The single hardest gate. Reviewers listen for a first-language-Arabic speaker's natural English, and imitated accents are screened out. Applicants are commonly asked to state their Arabic dialect background and where they learned English.
- Usable recording conditions. You do not need studio gear, but you do need a quiet space and a device whose microphone produces clean audio. Some screens ask for a short sample.
- Instruction-following in written English. The guide is text-only, in English, with no support call. You must be able to parse and execute it without clarification.
Logistics and pay
Fully remote and fully asynchronous — you record whenever you like within the turnaround window, which on quick-turnaround audio batches is usually a few days. Compensation is stated by the platform as a flat $50 for a single accepted recording submission; the $30/hr figure is a derived equivalent based on the estimated task length, not a guaranteed hourly rate, and no ongoing volume is promised. Treat this as a one-off with the possibility of being invited back for future batches if your audio is clean and your accent profile is in demand.