coqui/XTTS-v2
otherⓍTTS ⓍTTS is a Voice generation model that lets you clone voices into different languages by using just a quick 6-second audio clip. There i...
nari-labs/Dia-1.6B
apache-2.0Dia is a 1.6B parameter text to speech model created by Nari Labs. It was pushed to the Hub using the PytorchModelHubMixin( integration. Dia...
sesame/csm-1b
apache-2.0Visit HuggingFace for more details.
suno/bark
mitBark Bark is a transformer-based text-to-audio model created by Suno( Bark can generate highly realistic, multilingual speech as well as oth...
Zyphra/Zonos-v0.1-hybrid
apache-2.0Zonos-v0.1 --- Zonos-v0.1 is a leading open-weight text-to-speech model trained on more than 200k hours of varied multilingual speech, deliv...
SWivid/F5-TTS
cc-by-nc-4.0Download F5-TTS( or E2 TTS( and place under ckpts/ ckpts/ F5TTSv1Base/ model1250000.safetensors F5TTSBase/ model1200000.safetensors E2TTSBas...
canopylabs/orpheus-3b-0.1-ft
apache-2.0Visit HuggingFace for more details.
Kokoro-82M
apache-2.0Official text-to-speech model by hexgrad.
XTTS-v2
otherOfficial text-to-speech model by coqui.
chatterbox
mitOfficial text-to-speech model by ResembleAI.