nari-labs/Dia-1.6B
apache-2.0Dia is a 1.6B parameter text to speech model created by Nari Labs. It was pushed to the Hub using the PytorchModelHubMixin( integration. Dia...
sesame/csm-1b
apache-2.0Visit HuggingFace for more details.
suno/bark
mitBark Bark is a transformer-based text-to-audio model created by Suno( Bark can generate highly realistic, multilingual speech as well as oth...
Zyphra/Zonos-v0.1-hybrid
apache-2.0Zonos-v0.1 --- Zonos-v0.1 is a leading open-weight text-to-speech model trained on more than 200k hours of varied multilingual speech, deliv...
SWivid/F5-TTS
cc-by-nc-4.0Download F5-TTS( or E2 TTS( and place under ckpts/ ckpts/ F5TTSv1Base/ model1250000.safetensors F5TTSBase/ model1200000.safetensors E2TTSBas...
metavoiceio/metavoice-1B-v0.1
apache-2.0MetaVoice-1B is a 1.2B parameter base model trained on 100K hours of speech for TTS (text-to-speech). It has been built with the following p...
microsoft/speecht5_tts
mitSpeechT5 (TTS task) SpeechT5 model fine-tuned for speech synthesis (text-to-speech) on LibriTTS. This model was introduced in SpeechT5: Unif...
SparkAudio/Spark-TTS-0.5B
cc-by-nc-sa-4.0--- license: cc-by-nc-sa-4.0 language: - en - zh tags: - text-to-speech librarytag: spark-tts --- Spark-TTS Official model for Spark-TTS: An...
fishaudio/fish-speech-1.5
cc-by-nc-sa-4.0Fish Speech V1.5 Fish Speech V1.5 is a leading text-to-speech (TTS) model trained on more than 1 million hours of audio data in multiple lan...
hexgrad/Kokoro-82M
apache-2.0Kokoro is an open-weight TTS model with 82 million parameters. Despite its lightweight architecture, it delivers comparable quality to large...