parler-tts/parler-tts-large-v1

apache-2.0

Parler-TTS Large v1 Parler-TTS Large v1 is a 2.2B-parameters text-to-speech (TTS) model, trained on 45K hours of audio data, that can genera...

text to speechBy parler-tts

facebook/seamless-streaming

cc-by-nc-4.0

SeamlessStreaming SeamlessStreaming is a multilingual streaming translation model. It supports: - Streaming Automatic Speech Recognition on ...

text to speechBy facebook

WhisperSpeech/WhisperSpeech

mit

WhisperSpeech !Test it out yourself in Colab( !( If you have questions or you want to help you can find us in the \audio-generation channel ...

text to speechBy WhisperSpeech

espnet/kan-bayashi_ljspeech_vits

cc-by-4.0

ESPnet2 TTS pretrained model kan-bayashi/ljspeechvits ♻️ Imported from This model was trained by kan-bayashi using ljspeech/tts1 recipe in e...

text to speechBy espnet

suno/bark-small

mit

Bark Bark is a transformer-based text-to-audio model created by Suno( Bark can generate highly realistic, multilingual speech as well as oth...

text to speechBy suno

OuteAI/Llama-OuteTTS-1.0-1B

cc-by-nc-sa-4.0

Oute A I outeai.com Discord @OuteAI Llama OuteTTS 1.0 1B Llama OuteTTS 1.0 1B GGUF GitHub Library > !IMPORTANT > Important Sampling Consider...

text to speechBy OuteAI

facebook/mms-tts

cc-by-nc-4.0

Massively Multilingual Speech (MMS) : Text-to-Speech Models This repository contains a collection of text-to-speech (TTS) models, offering s...

text to speechBy facebook

canopylabs/orpheus-3b-0.1-pretrained

apache-2.0

Visit HuggingFace for more details.

text to speechBy canopylabs

hexgrad/Kokoro-82M

apache-2.0

Kokoro is an open-weight TTS model with 82 million parameters. Despite its lightweight architecture, it delivers comparable quality to large...

text to speechBy hexgrad

coqui/XTTS-v2

other

ⓍTTS ⓍTTS is a Voice generation model that lets you clone voices into different languages by using just a quick 6-second audio clip. There i...

text to speechBy coqui
Showing 10 of 55