Skip to main content
Echoryte

Reviewed language guide

Sundanese Audio to Text Transcription

Transcribe Sundanese audio into editable, word-timed text, then check spelling, names, numbers, boundaries, and mixed-language passages against the source.

Source profile

Language
Basa Sunda
BCP 47
su
Writing system
CLDR likely/default locale: Latn; not audio/sole script—confirm project choice
Direction
Left to right

Sundanese (Basa Sunda) review centers on this evidence: review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. For underspecified su, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Sundanese orthography. Sundanese projects may require Latin or Sundanese script. Orthographic targets include diacritics, word boundaries, proper names, number expressions, and code-switched terms. Project decision: whether Latin or Sundanese script and which regional register the audience needs. Before batch editing su, a reviewer must record whether Latin or Sundanese script and which regional register the audience needs and attach a source timestamp to disputed text.

Where this workflow fits

  • Sundanese interviews and oral histories
  • Sundanese lectures and training recordings
  • Source-linked Sundanese research review

Sundanese orthography and speech checkpoints

review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Editorial decision: whether Latin or Sundanese script and which regional register the audience needs. Visual inspection should cover record the project decision about whether Latin or Sundanese script and which regional register the audience needs.

Script choice and source-linked review for Basa Sunda

For underspecified su, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Sundanese orthography. Sundanese projects may require Latin or Sundanese script. For multilingual research sessions with productive affixes and reduplication, inspect review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Visually inspect diacritics, word boundaries, proper names, number expressions, and code-switched terms; record the project decision about whether Latin or Sundanese script and which regional register the audience needs. Replay a complete clause around every uncertain name, quantity, interruption, or ending.

Frozen deployment boundary

This page is published only because its language is present in capability snapshot m10-2026-07-20.1. It does not extend that catalog or claim a measured accuracy score.

  • The su hint maps to the reviewed su provider root and preserves su as the user's transcript language tag.
  • CLDR Latn is a likely/default locale hint, not an audio property; capability flags remain per language and model while editors confirm the project's real script and orthography.

Reviewed provider capability

assemblyai · universal-2

root su · tiers fast, standard, precision

Automatic detection
Reviewed no
Diarization
Reviewed no
Word timestamps
Reviewed yes

Questions from the workflow

What must a Sundanese editor decide before correcting the transcript?

affixes, repeated forms, particles, and borrowed vocabulary should be checked without forcing speech into a monolingual-looking draft. For Sundanese, review Sundanese affixes, reduplication, particles, speech levels, names, and Indonesian code-switching after script selection. Project choice: whether Latin or Sundanese script and which regional register the audience needs. Keep unresolved forms beside their source timestamps.

Reviewed sources