pyannote/speaker-diarization

mit

Visit HuggingFace for more details.

automatic speech recognitionBy pyannote

city96/FLUX.1-dev-gguf

other

This is a direct GGUF conversion of black-forest-labs/FLUX.1-dev( As this is a quantized model not a finetune, all the same restrictions/ori...

text to imageBy city96

microsoft/bitnet-b1.58-2B-4T

mit

BitNet b1.58 2B4T - Scaling Native 1-bit LLM This repository contains the weights for BitNet b1.58 2B4T, the first open-source, native 1-bit...

text generationBy microsoft

nvidia/parakeet-tdt-0.6b-v2

cc-by-4.0

🦜 Parakeet TDT 0.6B V2 (En) img { display: inline; } !Model architecture( | !Model size( | !Language( > 🎉 NEW: Multilingual Parakeet TDT 0...

automatic speech recognitionBy nvidia

SWivid/F5-TTS

cc-by-nc-4.0

Download F5-TTS( or E2 TTS( and place under ckpts/ ckpts/ F5TTSv1Base/ model1250000.safetensors F5TTSBase/ model1200000.safetensors E2TTSBas...

text to speechBy SWivid

dreamlike-art/dreamlike-diffusion-1.0

other

Dreamlike Diffusion 1.0 is SD 1.5 fine tuned on high quality art, made by dreamlike.art( If you want to use dreamlike models on your website...

text to imageBy dreamlike-art

ggerganov/whisper.cpp

mit

OpenAI's Whisper models converted to ggml format for use with whisper.cpp( Available models( | Model | Disk | SHA | | ------------------- | ...

automatic speech recognitionBy ggerganov

stabilityai/sv3d

other

Visit HuggingFace for more details.

image to videoBy stabilityai

comfyanonymous/flux_text_encoders

apache-2.0

Flux text encoder checkpoints meant to be used with the DualClipLoader node of ComfyUI( See the ComfyUI Flux examples(

GeneralBy comfyanonymous

Envvi/Inkpunk-Diffusion

creativeml-openrail-m

Inkpunk Diffusion Finetuned Stable Diffusion model trained on dreambooth. Vaguely inspired by Gorillaz, FLCL, and Yoji Shinkawa. Use nvinkpu...

text to imageBy Envvi
Showing 10 of 1932