parler-tts/parler-tts-large-v1
apache-2.0Parler-TTS Large v1 Parler-TTS Large v1 is a 2.2B-parameters text-to-speech (TTS) model, trained on 45K hours of audio data, that can genera...
facebook/seamless-streaming
cc-by-nc-4.0SeamlessStreaming SeamlessStreaming is a multilingual streaming translation model. It supports: - Streaming Automatic Speech Recognition on ...
WhisperSpeech/WhisperSpeech
mitWhisperSpeech !Test it out yourself in Colab( !( If you have questions or you want to help you can find us in the \audio-generation channel ...
espnet/kan-bayashi_ljspeech_vits
cc-by-4.0ESPnet2 TTS pretrained model kan-bayashi/ljspeechvits ♻️ Imported from This model was trained by kan-bayashi using ljspeech/tts1 recipe in e...
suno/bark-small
mitBark Bark is a transformer-based text-to-audio model created by Suno( Bark can generate highly realistic, multilingual speech as well as oth...
OuteAI/Llama-OuteTTS-1.0-1B
cc-by-nc-sa-4.0Oute A I outeai.com Discord @OuteAI Llama OuteTTS 1.0 1B Llama OuteTTS 1.0 1B GGUF GitHub Library > !IMPORTANT > Important Sampling Consider...
facebook/mms-tts
cc-by-nc-4.0Massively Multilingual Speech (MMS) : Text-to-Speech Models This repository contains a collection of text-to-speech (TTS) models, offering s...
canopylabs/orpheus-3b-0.1-pretrained
apache-2.0Visit HuggingFace for more details.
hexgrad/Kokoro-82M
apache-2.0Kokoro is an open-weight TTS model with 82 million parameters. Despite its lightweight architecture, it delivers comparable quality to large...
coqui/XTTS-v2
otherⓍTTS ⓍTTS is a Voice generation model that lets you clone voices into different languages by using just a quick 6-second audio clip. There i...