deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
mitDeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
google/gemma-7b-it
gemmaVisit HuggingFace for more details.
nvidia/parakeet-tdt-0.6b-v2
cc-by-4.0🦜 Parakeet TDT 0.6B V2 (En) img { display: inline; } !Model architecture( | !Model size( | !Language( > 🎉 NEW: Multilingual Parakeet TDT 0...
CompVis/stable-diffusion
creativeml-openrail-mStable Diffusion Stable Diffusion is a latent text-to-image diffusion model capable of generating photo-realistic images given any text inpu...
manycore-research/SpatialLM-Llama-1B
llama3.2SpatialLM-Llama-1B Introduction SpatialLM is a 3D large language model designed to process 3D point cloud data and generate structured 3D sc...
intfloat/multilingual-e5-large
mitMultilingual-E5-large Multilingual E5 Text Embeddings: A Technical Report( Liang Wang, Nan Yang, Xiaolong Huang, Linjun Yang, Rangan Majumde...
nlpconnect/vit-gpt2-image-captioning
apache-2.0nlpconnect/vit-gpt2-image-captioning This is an image captioning model trained by @ydshieh in flax ( this is pytorch version of this( The Il...
Qwen/Qwen3-30B-A3B
apache-2.0Qwen3-30B-A3B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of den...
google/gemma-2-2b-it
gemmaVisit HuggingFace for more details.
pyannote/speaker-diarization
mitVisit HuggingFace for more details.