deepseek-ai/DeepSeek-R1-Distill-Qwen-32B
mitDeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
mistralai/Mistral-Small-3.1-24B-Instruct-2503
apache-2.0Model Card for Mistral-Small-3.1-24B-Instruct-2503 Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art visi...
stabilityai/sd-vae-ft-mse-original
mitImproved Autoencoders Utilizing These weights are intended to be used with the original CompVis Stable Diffusion codebase( If you are lookin...
suno/bark
mitBark Bark is a transformer-based text-to-audio model created by Suno( Bark can generate highly realistic, multilingual speech as well as oth...
Salesforce/blip-image-captioning-large
bsd-3-clauseBLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...
microsoft/phi-1_5
mitModel Summary The language model Phi-1.5 is a Transformer with 1.3 billion parameters. It was trained using the same data sources as phi-1( ...
stabilityai/stable-cascade
otherStable Cascade This model is built upon the Würstchen( architecture and its main difference to other models like Stable Diffusion is that it...
01-ai/Yi-34B
apache-2.0Building the Next Generation of Open-Source and Bilingual LLMs 🤗 Hugging Face • 🤖 ModelScope • ✡️ WiseModel 👩🚀 Ask questions or discuss...
Wan-AI/Wan2.1-T2V-14B
apache-2.0Wan2.1 💜 Wan    |    🖥️ GitHub    |   🤗 Hugging Face   |   🤖 ModelScope   | &nbs...
TinyLlama/TinyLlama-1.1B-Chat-v1.0
apache-2.0TinyLlama-1.1B The TinyLlama project aims to pretrain a 1.1B Llama model on 3 trillion tokens. With some proper optimization, we can achieve...