Sundanese (Basa Sunda) review centers on this evidence: review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. For underspecified su, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Sundanese orthography. Sundanese projects may require Latin or Sundanese script. Orthographic targets include diacritics, word boundaries, proper names, number expressions, and code-switched terms. Project decision: whether Latin or Sundanese script and which regional register the audience needs. Before batch editing su, a reviewer must record whether Latin or Sundanese script and which regional register the audience needs and attach a source timestamp to disputed text.
Where this workflow fits
- Sundanese interviews and oral histories
- Sundanese lectures and training recordings
- Source-linked Sundanese research review
Sundanese orthography and speech checkpoints
review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Editorial decision: whether Latin or Sundanese script and which regional register the audience needs. Visual inspection should cover record the project decision about whether Latin or Sundanese script and which regional register the audience needs.
Script choice and source-linked review for Basa Sunda
For underspecified su, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Sundanese orthography. Sundanese projects may require Latin or Sundanese script. For multilingual research sessions with productive affixes and reduplication, inspect review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Visually inspect diacritics, word boundaries, proper names, number expressions, and code-switched terms; record the project decision about whether Latin or Sundanese script and which regional register the audience needs. Replay a complete clause around every uncertain name, quantity, interruption, or ending.
Frozen deployment boundary
This page is published only because its language is present in capability snapshot m10-2026-07-20.1. It does not extend that catalog or claim a measured accuracy score.
- The su hint maps to the reviewed su provider root and preserves su as the user's transcript language tag.
- CLDR Latn is a likely/default locale hint, not an audio property; capability flags remain per language and model while editors confirm the project's real script and orthography.
Reviewed provider capability
assemblyai · universal-2
root su · tiers fast, standard, precision
- Automatic detection
- Reviewed no
- Diarization
- Reviewed no
- Word timestamps
- Reviewed yes
Questions from the workflow
What must a Sundanese editor decide before correcting the transcript?
affixes, repeated forms, particles, and borrowed vocabulary should be checked without forcing speech into a monolingual-looking draft. For Sundanese, review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Project choice: whether Latin or Sundanese script and which regional register the audience needs. Keep unresolved forms beside their source timestamps.
Reviewed sources
- IANA Language Subtag Registry · checked 2026-07-20
- AssemblyAI model and supported-language documentation · checked 2026-07-20
- AssemblyAI pre-recorded audio supported languages · checked 2026-07-20
- Unicode CLDR likely-subtags language and script data · checked 2026-07-20
- Sundanese language and orthography editorial reference · checked 2026-07-20
- World Atlas of Language Structures Online · checked 2026-07-20