meta-llama/Llama-2-7b-chat
llama2Visit HuggingFace for more details.
cognitivecomputations/WizardLM-13B-Uncensored
otherThis is WizardLM trained with a subset of the dataset - responses that contained alignment / moralizing were removed. The intent is to train...
MiniMaxAI/MiniMax-Text-01
WeChat MiniMax-Text-01 1. Introduction MiniMax-Text-01 is a powerful language model with 456 billion total parameters, of which 45.9 billion...
tencent/Tencent-Hunyuan-Large
other GITHUB   |   🖥️  official website  |  🕖   HunyuanAPI|  🐳   Gitee Tec...
Qwen/Qwen3-30B-A3B
apache-2.0Qwen3-30B-A3B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of den...
TheBloke/Mistral-7B-Instruct-v0.1-GGUF
apache-2.0Chat & support: TheBloke's Discord server Want to contribute? TheBloke's Patreon page TheBloke's LLM work is generously supported by a grant...
meta-llama/Llama-3.2-3B
llama3.2Visit HuggingFace for more details.
deepseek-ai/DeepSeek-R1
mitDeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
meta-llama/Meta-Llama-3-8B
llama3Visit HuggingFace for more details.
bigscience/bloom
bigscience-bloom-rail-1.0BigScience Large Open-science Open-access Multilingual Language Model Version 1.3 / 6 July 2022 Current Checkpoint: Training Iteration 95000...