meta-llama/Llama-2-7b-hf
llama2Visit HuggingFace for more details.
Qwen/Qwen2.5-Coder-32B-Instruct
apache-2.0Qwen2.5-Coder-32B-Instruct Introduction Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as Co...
EleutherAI/gpt-j-6b
apache-2.0GPT-J 6B Model Description GPT-J 6B is a transformer model trained using Ben Wang's Mesh Transformer JAX( "GPT-J" refers to the class of mod...
deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
mitDeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
google/gemma-7b-it
gemmaVisit HuggingFace for more details.
manycore-research/SpatialLM-Llama-1B
llama3.2SpatialLM-Llama-1B Introduction SpatialLM is a 3D large language model designed to process 3D point cloud data and generate structured 3D sc...
Qwen/Qwen3-30B-A3B
apache-2.0Qwen3-30B-A3B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of den...
google/gemma-2-2b-it
gemmaVisit HuggingFace for more details.
meta-llama/Llama-2-70b-hf
llama2Visit HuggingFace for more details.
upstage/SOLAR-10.7B-Instruct-v1.0
cc-by-nc-4.0Meet 10.7B Solar: Elevating Performance with Upstage Depth UP Scaling! (This model is upstage/SOLAR-10.7B-v1.0( fine-tuned version for singl...