Dialect, code-switching, and short vowels absent from writing require listening rather than text-only correction. In the target track, Choose concise written phrasing, keep names consistent, and do not infer a spoken regional variety from Han text alone. This pair crosses text direction, so cue alignment, punctuation, numerals, and embedded names require a real bidirectional rendering check.
Decision signals
- Source language: Arabic (ar), written with Arabic.
- Target language: Simplified Chinese (zh-Hans), written with Han.
- The workflow preserves cue IDs and timing first, then makes visible text and line-break changes under human review.
Prepare the Arabic cue track
Cue breaks should preserve connected phrases and must be tested in an actual right-to-left subtitle renderer. Inspect bidirectional punctuation, numbers, Latin abbreviations, and bracket order in the rendered player. Correct uncertain names, numbers, and speaker changes against the recording before asking a translation model to transform the text.
Shape readable Simplified Chinese subtitles
Choose concise written phrasing, keep names consistent, and do not infer a spoken regional variety from Han text alone. Chinese lines do not use spaces as word boundaries, so break by clause and meaning rather than character count alone. The source review policy was different: Select a documented Modern Standard or regional policy and keep names and technical terms consistent. Keep a stable mapping back to the source cue so reviewers can compare meaning without losing the original timing context.
Run bilingual timing and rendering QA
Use full-width punctuation consistently and inspect Latin abbreviations, numbers, and units at script boundaries. This pair crosses text direction, so cue alignment, punctuation, numerals, and embedded names require a real bidirectional rendering check. On the source side, Dialect, code-switching, and short vowels absent from writing require listening rather than text-only correction. Watch the result with audio at normal speed, inspect every speaker change, and export only the reviewed track rather than an unexamined model response.
Boundaries to keep visible
- The page documents a Arabic-to-Simplified Chinese review workflow; it is not an accuracy score or a promise that every configured model, account, or deployment accepts the pair.
- Keep source cue timestamps as the initial anchor, but permit a human subtitle editor to revise segmentation when target-language readability would otherwise fail.
Reviewed sources
- W3C WebVTT specification checked 2026-07-19
- Unicode bidirectional algorithm checked 2026-07-19
- Arabic language and orthography editorial reference checked 2026-07-20
- Mandarin Chinese language and orthography editorial reference checked 2026-07-20