fishaudio/fish-speech-1.5
cc-by-nc-sa-4.0Fish Speech V1.5 Fish Speech V1.5 is a leading text-to-speech (TTS) model trained on more than 1 million hours of audio data in multiple lan...
HKUSTAudio/Llasa-3B
cc-by-nc-4.0!arXiv( Update (2025-05-10): Sometimes I find that topp=0.95 and temperature=0.9 produce more stable results. Update (2025-02-13): Add Llasa...
myshell-ai/OpenVoiceV2
mitOpenVoice V2 In April 2024, we release OpenVoice V2, which includes all features in V1 and has: 1. Better Audio Quality. OpenVoice V2 adopts...
Zyphra/Zonos-v0.1-transformer
apache-2.0Zonos-v0.1 --- Zonos-v0.1 is a leading open-weight text-to-speech model trained on more than 200k hours of varied multilingual speech, deliv...
ByteDance/MegaTTS3
apache-2.0Model Description This is a huggingface model card for MegaTTS 3 👋 - Paper: MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transform...
lj1995/GPT-SoVITS
mitpretrained models used in
moonshotai/Kimi-Audio-7B-Instruct
mitKimi-Audio 🤗 Kimi-Audio-7B | 🤗 Kimi-Audio-7B-Instruct | 📑 Paper Introduction We present Kimi-Audio, an open-source audio fou...
OuteAI/OuteTTS-0.2-500M
cc-by-nc-4.0table { border-collapse: collapse; width: 100%; margin-bottom: 20px; } th, td { border: 1px solid ddd; padding: 8px; text-align: center; } ....
myshell-ai/MeloTTS-English
mitMeloTTS MeloTTS is a high-quality multi-lingual text-to-speech library by MIT( and MyShell.ai( Supported languages include: | Model card | E...
facebook/fastspeech2-en-ljspeech
fastspeech2-en-ljspeech FastSpeech 2( text-to-speech model from fairseq S^2 (paper( - English - Single-speaker female voice - Trained on LJS...