Skip to main content

← All model recipe tasks

Text-to-Speech

Models that synthesize audio from text.

Classification follows the Hugging Face task taxonomy. See the Hugging Face Text-to-Speech task page for the ecosystem-level task definition.

Model families

Model familyDeclared recipesExact checkpoint examplesRuntime CLI
bark5
suno/bark
suno/bark-small
trtmc generate-audio
magpie_tts2
nvidia/magpie_tts_multilingual_357m
trtmc generate-audio