Fast conversational English can compress unstressed words, so replay names and negatives at normal speed. In the target track, Keep grammatical gender, formality, and regional vocabulary consistent across the whole subtitle track. Both writing systems render left to right, but player layout still needs a visual pass.
Decision signals
- Source language: English (en), written with Latin.
- Target language: Spanish (es), written with Latin.
- The workflow preserves cue IDs and timing first, then makes visible text and line-break changes under human review.
Prepare the English cue track
English cue breaks usually read best at phrase boundaries, not immediately before short function words. Review apostrophes, quotation style, contractions, and sentence case after line wrapping. Correct uncertain names, numbers, and speaker changes against the recording before asking a translation model to transform the text.
Shape readable Spanish subtitles
Keep grammatical gender, formality, and regional vocabulary consistent across the whole subtitle track. Spanish articles and clitic pronouns should remain attached to a readable phrase when cues split. The source review policy was different: Prefer direct clauses and preserve intentional register rather than expanding every implicit subject. Keep a stable mapping back to the source cue so reviewers can compare meaning without losing the original timing context.
Run bilingual timing and rendering QA
Preserve opening question and exclamation marks and check dialogue dashes deliberately. Both writing systems render left to right, but player layout still needs a visual pass. On the source side, Fast conversational English can compress unstressed words, so replay names and negatives at normal speed. Watch the result with audio at normal speed, inspect every speaker change, and export only the reviewed track rather than an unexamined model response.
Boundaries to keep visible
- The page documents a English-to-Spanish review workflow; it is not an accuracy score or a promise that every configured model, account, or deployment accepts the pair.
- Keep source cue timestamps as the initial anchor, but permit a human subtitle editor to revise segmentation when target-language readability would otherwise fail.
Reviewed sources
- W3C WebVTT specification checked 2026-07-19
- Unicode bidirectional algorithm checked 2026-07-19
- American English language and orthography editorial reference checked 2026-07-20
- British English language and orthography editorial reference checked 2026-07-20
- Spanish language and orthography editorial reference checked 2026-07-20