meta-llama/Llama-3.1-405B
llama3.1Visit HuggingFace for more details.
mistralai/Mistral-Small-24B-Instruct-2501
apache-2.0Model Card for Mistral-Small-24B-Instruct-2501 Mistral Small 3 ( 2501 ) sets a new benchmark in the "small" Large Language Models category b...
deepseek-ai/DeepSeek-R1-Zero
mitDeepSeek-R1 Paper LinkποΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
Qwen/Qwen3-235B-A22B
apache-2.0Qwen3-235B-A22B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of d...
microsoft/Phi-3.5-mini-instruct
mitπPhi-4: multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary Phi-3.5-mini is a lightweight, state-of-the-art open model built ...
meta-llama/Meta-Llama-3-70B
llama3Visit HuggingFace for more details.
meta-llama/Llama-2-70b-hf
llama2Visit HuggingFace for more details.
Qwen/Qwen2.5-72B-Instruct
otherQwen2.5-72B-Instruct Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base langu...
WizardLMTeam/WizardCoder-15B-V1.0
bigscience-openrail-mWizardCoder: Empowering Code Large Language Models with Evol-Instruct π Home Page π€ HF Repo β’π± Github Repo β’ π¦ Twitter π WizardLM β’ π ...
meta-llama/Llama-3.1-70B-Instruct
llama3.1Visit HuggingFace for more details.