Automatic Speech Recognition
Speech-to-text models, including offline and streaming contracts.
Classification follows the Hugging Face task taxonomy. See the Hugging Face Automatic Speech Recognition task page for the ecosystem-level task definition.
Model families
| Model family | Declared recipes | Exact checkpoint examples | Runtime CLI |
|---|---|---|---|
canary | 2 | nvidia/canary-1b-v2 | trtmc transcribe |
nemotron_speech_streaming | 3 | nvidia/nemotron-3.5-asr-streaming-0.6bnvidia/nemotron-speech-streaming-en-0.6b | trtmc transcribe |
whisper | 4 | openai/whisper-large-v3-turboopenai/whisper-tiny | trtmc transcribe |