USE WHEN
Use this manual when
- Chinese-to-English localization or the reverse
- Multi-speaker productions with names and retained music/effects
- Teams delivering MP4, WAV, SRT or AAF packages
PRODUCTION MANUAL · Audio & localization
Unify terminology, voice rights, speakers, timecodes, captions and stems in one dubbing QC process instead of publishing automatic output unchecked.

USE WHEN
DO NOT USE AS
VERIFIED FACTS
ElevenLabs currently documents 90+ languages, speaker separation and retained background audio; Automatic Dubbing v2 remains labelled Alpha and recommends no more than nine unique speakers for best results.
ElevenLabs documents exports including MP4, AAC, AAF, SRT and separated WAV tracks, allowing editable delivery instead of a single flattened video.
YouTube warns that automatic dubs can fail on proper nouns, idioms, accents, dialects, background noise and speaker matching. Auto dubs are not directly editable, so target-language review should precede publication.
REQUIRED INPUTS
WORKFLOW
Verify source dialogue, speakers and timecodes before translation.
Build a bilingual glossary and flag names, wordplay and cultural references for human work.
Turn literal translation into performable dialogue while preserving meaning, relationships and duration.
Generate test segments only with authorized voices and check timbre, accent, emotion and separation.
Review meaning, pronunciation, pacing, lip tolerance and background-audio damage line by line.
Export video, dialogue stems, captions, glossary and revision log with a path back to the source master.
DELIVERABLES & QC
EVIDENCE BOUNDARY
Language count, separation and export formats come from official documentation and do not guarantee broadcast quality for every language, accent or mix. Similarity does not replace voice rights; final dubs require target-language and audio QC.