meta-llama/Llama-3.2-1B
llama3.2Visit HuggingFace for more details.
Qwen/Qwen2.5-Coder-32B-Instruct
apache-2.0Qwen2.5-Coder-32B-Instruct Introduction Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as Co...
mistralai/Mistral-7B-Instruct-v0.3
apache-2.0Model Card for Mistral-7B-Instruct-v0.3 The Mistral-7B-Instruct-v0.3 Large Language Model (LLM) is an instruct fine-tuned version of the Mis...
Qwen/QwQ-32B-Preview
apache-2.0QwQ-32B-Preview Introduction QwQ-32B-Preview is an experimental research model developed by the Qwen Team, focused on advancing AI reasoning...
CohereLabs/c4ai-command-r-plus
cc-by-nc-4.0Visit HuggingFace for more details.
HuggingFaceH4/zephyr-7b-beta
mitModel Card for Zephyr 7B β Zephyr is a series of language models that are trained to act as helpful assistants. Zephyr-7B-β is the second mo...
mistralai/Mixtral-8x7B-v0.1
apache-2.0Model Card for Mixtral-8x7B The Mixtral-8x7B Large Language Model (LLM) is a pretrained generative Sparse Mixture of Experts. The Mistral-8x...
mistralai/Mistral-7B-Instruct-v0.1
apache-2.0Model Card for Mistral-7B-Instruct-v0.1 Encode and Decode with mistralcommon py from mistralcommon.tokens.tokenizers.mistral import MistralT...
meta-llama/Llama-3.1-8B
llama3.1Visit HuggingFace for more details.
EleutherAI/gpt-j-6b
apache-2.0GPT-J 6B Model Description GPT-J 6B is a transformer model trained using Ben Wang's Mesh Transformer JAX( "GPT-J" refers to the class of mod...