microsoft/Phi-3-vision-128k-instruct

mit

🎉 Phi-3.5: mini-instruct( MoE-instruct( ; vision-instruct( Model Summary The Phi-3-Vision-128K-Instruct is a lightweight, state-of-the-art ...

text generationBy microsoft

meta-llama/Llama-3.2-1B-Instruct

llama3.2

Visit HuggingFace for more details.

text generationBy meta-llama

meta-llama/Llama-3.1-405B

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama

mistralai/Mistral-Small-24B-Instruct-2501

apache-2.0

Model Card for Mistral-Small-24B-Instruct-2501 Mistral Small 3 ( 2501 ) sets a new benchmark in the "small" Large Language Models category b...

text generationBy mistralai

deepseek-ai/DeepSeek-R1-Zero

mit

DeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

Qwen/Qwen3-235B-A22B

apache-2.0

Qwen3-235B-A22B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of d...

text generationBy Qwen

microsoft/Phi-3.5-mini-instruct

mit

🎉Phi-4: multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary Phi-3.5-mini is a lightweight, state-of-the-art open model built ...

text generationBy microsoft

meta-llama/Meta-Llama-3-70B

llama3

Visit HuggingFace for more details.

text generationBy meta-llama

tencent/Tencent-Hunyuan-Large

other

&nbspGITHUB&nbsp&nbsp | &nbsp&nbsp🖥️&nbsp&nbspofficial website&nbsp&nbsp|&nbsp&nbsp🕖&nbsp&nbsp HunyuanAPI|&nbsp&nbsp🐳&nbsp&nbsp Gitee Tec...

text generationBy tencent

Qwen/Qwen2.5-72B-Instruct

other

Qwen2.5-72B-Instruct Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base langu...

text generationBy Qwen
Showing 10 of 563