meta-llama/Llama-3.1-70B-Instruct

llama3.1

Visit HuggingFace for more details.

text generationBy meta-llama

deepseek-ai/DeepSeek-Prover-V2-671B

1. Introduction We introduce DeepSeek-Prover-V2, an open-source large language model designed for formal theorem proving in Lean 4, with ini...

text generationBy deepseek-ai

deepseek-ai/DeepSeek-R1-Distill-Qwen-7B

mit

DeepSeek-R1 Paper LinkπŸ‘οΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

vikhyatk/moondream1

πŸŒ” moondream1 1.6B parameter model built by @vikhyatk( using SigLIP, Phi-1.5 and the LLaVa training dataset. The model is release for resear...

text generationBy vikhyatk

WizardLMTeam/WizardCoder-15B-V1.0

bigscience-openrail-m

WizardCoder: Empowering Code Large Language Models with Evol-Instruct 🏠 Home Page πŸ€— HF Repo β€’πŸ± Github Repo β€’ 🐦 Twitter πŸ“ƒ WizardLM β€’ πŸ“ƒ ...

text generationBy WizardLMTeam

PygmalionAI/pygmalion-6b

creativeml-openrail-m

Pygmalion 6B Model description Pymalion 6B is a proof-of-concept dialogue model based on EleutherAI's GPT-J-6B( Warning: This model is NOT s...

text generationBy PygmalionAI

google/gemma-2b-it

gemma

Visit HuggingFace for more details.

text generationBy google

Gustavosta/MagicPrompt-Stable-Diffusion

mit

MagicPrompt - Stable Diffusion This is a model from the MagicPrompt series of models, which are GPT-2( models intended to generate prompt te...

text generationBy Gustavosta

deepseek-ai/DeepSeek-R1-Distill-Llama-8B

mit

DeepSeek-R1 Paper LinkπŸ‘οΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...

text generationBy deepseek-ai

mistralai/Mixtral-8x22B-Instruct-v0.1

apache-2.0

Model Card for Mixtral-8x22B-Instruct-v0.1 Encode and Decode with mistralcommon py from mistralcommon.tokens.tokenizers.mistral import Mistr...

text generationBy mistralai
Showing 10 of 563