Qwen/Qwen2.5-7B-Instruct
apache-2.0Qwen2.5-7B-Instruct Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base langua...
nvidia/Nemotron-4-340B-Instruct
otherNemotron-4-340B-Instruct !Model architecture( size( Model Overview Nemotron-4-340B-Instruct is a large language model (LLM) that can be used...
deepseek-ai/DeepSeek-R1-Distill-Llama-70B
mitDeepSeek-R1 Paper LinkποΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
cardiffnlp/twitter-roberta-base-sentiment-latest
Twitter-roBERTa-base for Sentiment Analysis - UPDATED (2022) This is a RoBERTa-base model trained on ~124M tweets from January 2018 to Decem...
mixedbread-ai/mxbai-embed-large-v1
apache-2.0The crispy sentence embedding family from Mixedbread. π Looking for a simple end-to-end retrieval solution? Meet Omni, our multimodal and m...
stabilityai/stable-diffusion-2-1-base
openrail++Visit HuggingFace for more details.
XLabs-AI/flux-controlnet-collections
other!Controlnet collections for Flux( ( This repository provides a collection of ControlNet checkpoints for FLUX.1-dev model( by Black Forest La...
google/gemma-2-9b
gemmaVisit HuggingFace for more details.
allenai/olmOCR-7B-0225-preview
apache-2.0olmOCR-7B-0225-preview This is a preview release of the olmOCR model that's fine tuned from Qwen2-VL-7B-Instruct using the olmOCR-mix-0225( ...
agentica-org/DeepCoder-14B-Preview
mitDeepCoder-14B-Preview π Democratizing Reinforcement Learning for LLMs (RLLM) π DeepCoder Overview DeepCoder-14B-Preview is a code reasonin...