Mercor · Voice & audio
Audiobook QA Expert — English (US) - review AI outputs in your specialty
Listed on Mercor as “Audiobook QA Expert - English (US)”
What this actually is
You record voice samples following scripts, evaluate AI-generated audio for quality, and provide expert feedback used to train speech models. Often paid by the minute or hour of recording. The platform title (Audiobook QA Expert - English (US)) reflects the rate band and the expertise required, not the day-to-day work.
Advertisement
Can you do this on your visa?
F-2 / F-4 / F-5 / F-6: open. E-1 to E-7: needs concurrent-employment permit. D-2 / D-4 students: S-3 permit, 20 hr/week cap. D-10 / D-8: case by case.
Korean tax on USD income
First 5 years in Korea: foreign-source income only taxed if remitted into Korea. After year 5: worldwide income. Full tax guide.
Original posting from Mercor
Help make the next generation of AI-narrated audiobooks sound genuinely human.
1. Overview
We're building a team of native English speakers who love audiobooks to help evaluate AI-narrated (text-to-speech) audiobooks in US / North American English. You'll listen the way a real audiobook fan does - but with a careful, analytical ear - and flag the moments where the narration slips, so the underlying models can be made better.
This is remote, part-time work at roughly 20 hours per week (about 4 hours per day). You can be based anywhere, though we prefer people living in a region where English is natively spoken.
2. What You'll Do
- Listen to AI-narrated audiobooks in English and assess how natural, accurate, and enjoyable the narration is.
- Spot and log errors in the synthesized speech - skipped, added, or mispronounced words; numbers read incorrectly; abbreviations expanded wrongly; unnatural phrasing or intonation; audio glitches and artifacts.
- Categorize each issue and mark the exact span where it occurs, using our review tooling.
- Report on overall listener satisfaction and where the narration feels natural versus where it breaks the experience.
- Work consistently and carefully across many hours of audio without losing attention to detail.
3. Who We're Looking For
- Native or near-native fluency in English (US / North American English), and an active, genuine audiobook listener - someone who listens regularly and has real opinions about what makes a good listen.
- Enough working English to report clearly and communicate with the team.
- Some experience with annotation, labeling, transcription, proofreading, linguistics, or other detail-oriented review work.
- A sharp, patient ear and the discipline to stay accurate over long listening sessions.
- Reliable availability of about 20 hours per week on a part-time basis.
Nice to Have
- Prior experience on AI training-data, model-evaluation, or audio/QA annotation projects.
- A background in language, editing, voice, or audio production.
Quoted from Mercor’s public listing on 2026-09-15. We don’t edit platform copy; honest framing is in the title and the “what this actually is” block above.
Related AI training jobs
Alignerr · Voice & audio
Audio Transcript Editor - Portuguese Brasileiro (AI Training)
$10-$35/hr · Remote · USD
Alignerr · Voice & audio
Brazilian Portuguese Audio Transcript Editor (AI Training)
$10-$35/hr · Remote · USD
Alignerr · Voice & audio
Brazilian Portuguese Audio Transcript Editor - Remote (AI Training)
$10-$35/hr · Remote · USD
Alignerr · Voice & audio
Brazilian Portuguese Audio Transcriptionist (AI Training)
$10-$35/hr · Remote · USD
More on this platform
About Mercor
AI-interview-based talent network. One application, voice interview with their AI, then matched to projects across coding, research, and specialist work. Pay scales with track and seniority.
Mercor review: AI-interview talent network
4.1/5 on Glassdoor, fastest-growing platform in the category (+509% YoY). What the AI video interview actually asks, real pay across coding/research/medical/legal/finance tracks ($25-$200/hr), and the project-availability problem.
See all AI training jobs
Browse by category and compare across all eight platforms we cover.