nvidia/Llama-3.1-Nemotron-70B-Instruct-HF
llama3.1Model Overview Description: Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM...
meta-llama/Llama-3.2-1B
llama3.2Visit HuggingFace for more details.
Qwen/Qwen2.5-Coder-32B-Instruct
apache-2.0Qwen2.5-Coder-32B-Instruct Introduction Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as Co...
mistralai/Mistral-7B-Instruct-v0.3
apache-2.0Model Card for Mistral-7B-Instruct-v0.3 The Mistral-7B-Instruct-v0.3 Large Language Model (LLM) is an instruct fine-tuned version of the Mis...
Qwen/QwQ-32B-Preview
apache-2.0QwQ-32B-Preview Introduction QwQ-32B-Preview is an experimental research model developed by the Qwen Team, focused on advancing AI reasoning...
mistralai/Mistral-7B-Instruct-v0.1
apache-2.0Model Card for Mistral-7B-Instruct-v0.1 Encode and Decode with mistralcommon py from mistralcommon.tokens.tokenizers.mistral import MistralT...
microsoft/Phi-3-mini-128k-instruct
mit🎉Phi-4: multimodal-instruct( | onnx( mini-instruct( | onnx( Model Summary The Phi-3-Mini-128K-Instruct is a 3.8 billion-parameter, lightwei...
meta-llama/Llama-3.1-8B
llama3.1Visit HuggingFace for more details.
EleutherAI/gpt-j-6b
apache-2.0GPT-J 6B Model Description GPT-J 6B is a transformer model trained using Ben Wang's Mesh Transformer JAX( "GPT-J" refers to the class of mod...
meta-llama/Meta-Llama-3-70B-Instruct
llama3Visit HuggingFace for more details.