meta-llama/Llama-3.1-405B

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama

mistralai/Mistral-Small-24B-Instruct-2501

apache-2.0

Model Card for Mistral-Small-24B-Instruct-2501 Mistral Small 3 ( 2501 ) sets a new benchmark in the "small" Large Language Models category b...

text generationBy mistralai

deepseek-ai/DeepSeek-R1-Zero

mit

DeepSeek-R1 Paper LinkπŸ‘οΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

Qwen/Qwen3-235B-A22B

apache-2.0

Qwen3-235B-A22B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of d...

text generationBy Qwen

microsoft/Phi-3.5-mini-instruct

mit

πŸŽ‰Phi-4: multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary Phi-3.5-mini is a lightweight, state-of-the-art open model built ...

text generationBy microsoft

meta-llama/Meta-Llama-3-70B

llama3

Visit HuggingFace for more details.

text generationBy meta-llama

meta-llama/Llama-2-70b-hf

llama2

Visit HuggingFace for more details.

text generationBy meta-llama

Qwen/Qwen2.5-72B-Instruct

other

Qwen2.5-72B-Instruct Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base langu...

text generationBy Qwen

WizardLMTeam/WizardCoder-15B-V1.0

bigscience-openrail-m

WizardCoder: Empowering Code Large Language Models with Evol-Instruct 🏠 Home Page πŸ€— HF Repo β€’πŸ± Github Repo β€’ 🐦 Twitter πŸ“ƒ WizardLM β€’ πŸ“ƒ ...

text generationBy WizardLMTeam

meta-llama/Llama-3.1-70B-Instruct

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama
Showing 10 of 563