Languages
To choose a transcription model, refer to Models.
The languages, packs, and options on this page apply to the Enhanced and Standard models. The Melia 1 and Oak 1 models are multilingual: they transcribe the individual languages listed here and switch between them automatically, without language selection. You can use their language codes as language hints. Melia 1 and Oak 1 do not support the auto option, the bilingual and multi-language pack codes, or translation. For Melia 1 and Oak 1, refer to Models.
The Agent STT API uses the Linden 1 model, which supports the same languages as Realtime transcription. Translation is not available. For Agent STT, refer to Models.
Transcription languages
To automatically identify the language in an audio file, use the Language Identification feature.
To dynamically update your system with the latest languages and features offered by Speechmatics, use the Feature Discovery endpoint.
Speechmatics supports the following languages. Your ability to use any or all of them depends on the languages you are contracted to use.
Speechmatics takes a global-first approach to languages. A single language pack supports many accents and dialects, so you do not need to know which accent is in your audio before selecting a language. This approach achieves high accuracy compared to accent-specific language packs.
Each language is uniquely identified by a two-letter code (ISO 639-1) or three-letter code (ISO 639-3) in API requests and responses.
Medical languages
Speechmatics offers two models tuned for healthcare audio: the Enhanced Medical model, which transcribes a single selected language, and Oak 1, which is multilingual and detects the language automatically.
For a language outside this table, Enhanced Medical falls back to the standard Enhanced model, which still delivers high accuracy on healthcare audio without the medical tuning. Oak 1 transcribes all the individual languages listed above regardless of this table — the table lists only the languages where Oak 1 has the additional medical terminology uplift (procedures, medications, conditions, and anatomy); outside it, Oak 1 still transcribes the language, without that uplift.
Translation languages
Translation is available with the Enhanced and Standard models. It is supported for most Speechmatics languages, with the supported translation pairs listed below. For more details, see Translation.
Bilingual and multi-language packs
These packs handle a fixed set of languages that you select in advance. To transcribe audio without selecting languages, including spontaneous switching across all supported languages, use the Melia 1 multilingual model, or Oak 1 for healthcare audio. Refer to Models.
The Enhanced and Standard models can transcribe a selected combination of languages in one media file or stream, including speakers who switch between the languages in that pack. Each pack covers a fixed set of languages that you select with the language property.
Supported packs are:
This config selects the Mandarin and English pack:
{
"type": "transcription",
"transcription_config": {
"model": "enhanced",
"language": "cmn_en"
}
}
This config selects the Spanish and English pack, which requires the domain property:
{
"type": "transcription",
"transcription_config": {
"model": "enhanced",
"language": "es",
"domain": "bilingual-en"
}
}