meta-llama/Llama-2-7b-hf

llama2

Visit HuggingFace for more details.

text generationBy meta-llama

Qwen/Qwen2.5-Coder-32B-Instruct

apache-2.0

Qwen2.5-Coder-32B-Instruct Introduction Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as Co...

text generationBy Qwen

EleutherAI/gpt-j-6b

apache-2.0

GPT-J 6B Model Description GPT-J 6B is a transformer model trained using Ben Wang's Mesh Transformer JAX( "GPT-J" refers to the class of mod...

text generationBy EleutherAI

deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B

mit

DeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

google/gemma-7b-it

gemma

Visit HuggingFace for more details.

text generationBy google

manycore-research/SpatialLM-Llama-1B

llama3.2

SpatialLM-Llama-1B Introduction SpatialLM is a 3D large language model designed to process 3D point cloud data and generate structured 3D sc...

text generationBy manycore-research

Qwen/Qwen3-30B-A3B

apache-2.0

Qwen3-30B-A3B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of den...

text generationBy Qwen

google/gemma-2-2b-it

gemma

Visit HuggingFace for more details.

text generationBy google

meta-llama/Llama-2-70b-hf

llama2

Visit HuggingFace for more details.

text generationBy meta-llama

upstage/SOLAR-10.7B-Instruct-v1.0

cc-by-nc-4.0

Meet 10.7B Solar: Elevating Performance with Upstage Depth UP Scaling! (This model is upstage/SOLAR-10.7B-v1.0( fine-tuned version for singl...

text generationBy upstage
Showing 10 of 563