Text-to-Speech
Models that synthesize audio from text.
Classification follows the Hugging Face task taxonomy. See the Hugging Face Text-to-Speech task page for the ecosystem-level task definition.
Model families
| Model family | Declared recipes | Exact checkpoint examples | Runtime CLI |
|---|---|---|---|
bark | 5 | suno/barksuno/bark-small | trtmc generate-audio |
magpie_tts | 2 | nvidia/magpie_tts_multilingual_357m | trtmc generate-audio |