nvidia/Llama-3.1-Nemotron-70B-Instruct-HF

llama3.1

Model Overview Description: Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM...

text generationBy nvidia

meta-llama/Llama-3.2-1B

llama3.2

Visit HuggingFace for more details.

text generationBy meta-llama

Qwen/Qwen2.5-Coder-32B-Instruct

apache-2.0

Qwen2.5-Coder-32B-Instruct Introduction Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as Co...

text generationBy Qwen

mistralai/Mistral-7B-Instruct-v0.3

apache-2.0

Model Card for Mistral-7B-Instruct-v0.3 The Mistral-7B-Instruct-v0.3 Large Language Model (LLM) is an instruct fine-tuned version of the Mis...

text generationBy mistralai

Qwen/QwQ-32B-Preview

apache-2.0

QwQ-32B-Preview Introduction QwQ-32B-Preview is an experimental research model developed by the Qwen Team, focused on advancing AI reasoning...

text generationBy Qwen

mistralai/Mistral-7B-Instruct-v0.1

apache-2.0

Model Card for Mistral-7B-Instruct-v0.1 Encode and Decode with mistralcommon py from mistralcommon.tokens.tokenizers.mistral import MistralT...

text generationBy mistralai

microsoft/Phi-3-mini-128k-instruct

mit

🎉Phi-4: multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary The Phi-3-Mini-128K-Instruct is a 3.8 billion-parameter, lightwei...

text generationBy microsoft

meta-llama/Llama-3.1-8B

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama

EleutherAI/gpt-j-6b

apache-2.0

GPT-J 6B Model Description GPT-J 6B is a transformer model trained using Ben Wang's Mesh Transformer JAX( "GPT-J" refers to the class of mod...

text generationBy EleutherAI

meta-llama/Meta-Llama-3-70B-Instruct

llama3

Visit HuggingFace for more details.

text generationBy meta-llama
Showing 10 of 563