Estonian (eesti) review centers on this evidence: inspect quantity-sensitive forms, case endings, compounds, õ/ä/ö/ü, and words whose boundary changes grammatical analysis. For underspecified et, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Estonian orthography. Confirm the actual orthography and script with the project owner before normalizing any text. Orthographic targets include diacritics, word boundaries, proper names, number expressions, and code-switched terms. Project decision: how conversational forms and foreign names are represented in standard Estonian. Before batch editing et, a reviewer must record how conversational forms and foreign names are represented in standard Estonian and attach a source timestamp to disputed text.
Where this workflow fits
- Estonian interviews and oral histories
- Estonian lectures and training recordings
- Source-linked Estonian research review
Estonian orthography and speech checkpoints
inspect quantity-sensitive forms, case endings, compounds, õ/ä/ö/ü, and words whose boundary changes grammatical analysis. Editorial decision: how conversational forms and foreign names are represented in standard Estonian. Visual inspection should cover record the project decision about how conversational forms and foreign names are represented in standard Estonian.
Script choice and source-linked review for eesti
For underspecified et, CLDR supplies Latn (Latin script) as a likely/default web-locale hint—not an audio attribute or exclusive Estonian orthography. Confirm the actual orthography and script with the project owner before normalizing any text. For fieldwork or lectures with extensive case marking and compounding, inspect inspect quantity-sensitive forms, case endings, compounds, õ/ä/ö/ü, and words whose boundary changes grammatical analysis. Visually inspect diacritics, word boundaries, proper names, number expressions, and code-switched terms; record the project decision about how conversational forms and foreign names are represented in standard Estonian. Replay a complete clause around every uncertain name, quantity, interruption, or ending.
Frozen deployment boundary
This page is published only because its language is present in capability snapshot m10-2026-07-20.1. It does not extend that catalog or claim a measured accuracy score.
- The et hint maps to the reviewed et provider root and preserves et as the user's transcript language tag.
- CLDR Latn is a likely/default locale hint, not an audio property; capability flags remain per language and model while editors confirm the project's real script and orthography.
Reviewed provider capability
assemblyai · universal-2
root et · tiers fast, standard, precision
- Automatic detection
- Reviewed no
- Diarization
- Reviewed no
- Word timestamps
- Reviewed yes
Questions from the workflow
What must a Estonian editor decide before correcting the transcript?
long inflected words, vowel or consonant quantity, compounds, and names require full-sentence evidence rather than surface plausibility. For Estonian, inspect quantity-sensitive forms, case endings, compounds, õ/ä/ö/ü, and words whose boundary changes grammatical analysis. Project choice: how conversational forms and foreign names are represented in standard Estonian. Keep unresolved forms beside their source timestamps.
Reviewed sources
- IANA Language Subtag Registry · checked 2026-07-20
- AssemblyAI model and supported-language documentation · checked 2026-07-20
- AssemblyAI pre-recorded audio supported languages · checked 2026-07-20
- Unicode CLDR likely-subtags language and script data · checked 2026-07-20
- Estonian language and orthography editorial reference · checked 2026-07-20
- World Atlas of Language Structures Online · checked 2026-07-20