mistralai/Mixtral-8x7B-Instruct-v0.1

apache-2.0

Model Card for Mixtral-8x7B Tokenization with mistral-common py from mistralcommon.tokens.tokenizers.mistral import MistralTokenizer from mi...

text generationBy mistralai

bigscience/bloom

bigscience-bloom-rail-1.0

BigScience Large Open-science Open-access Multilingual Language Model Version 1.3 / 6 July 2022 Current Checkpoint: Training Iteration 95000...

text generationBy bigscience

meta-llama/Meta-Llama-3-8B

llama3

Visit HuggingFace for more details.

text generationBy meta-llama

meta-llama/Llama-2-7b-chat-hf

llama2

Visit HuggingFace for more details.

text generationBy meta-llama

meta-llama/Llama-2-7b

llama2

Visit HuggingFace for more details.

text generationBy meta-llama

meta-llama/Llama-3.1-8B-Instruct

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama

deepseek-ai/DeepSeek-R1

mit

DeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

meta-llama/Meta-Llama-3-8B-Instruct

llama3

Visit HuggingFace for more details.

text generationBy meta-llama

deepseek-ai/DeepSeek-V3

Paper Link👁️ 1. Introduction We present DeepSeek-V3, a strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B a...

text generationBy deepseek-ai

mistralai/Mistral-7B-v0.1

apache-2.0

Model Card for Mistral-7B-v0.1 The Mistral-7B-v0.1 Large Language Model (LLM) is a pretrained generative text model with 7 billion parameter...

text generationBy mistralai
Showing 10 of 563