meta-llama/Llama-3.1-70B-Instruct
llama3.1Visit HuggingFace for more details.
deepseek-ai/DeepSeek-Prover-V2-671B
1. Introduction We introduce DeepSeek-Prover-V2, an open-source large language model designed for formal theorem proving in Lean 4, with ini...
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
mitDeepSeek-R1 Paper LinkποΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
vikhyatk/moondream1
π moondream1 1.6B parameter model built by @vikhyatk( using SigLIP, Phi-1.5 and the LLaVa training dataset. The model is release for resear...
WizardLMTeam/WizardCoder-15B-V1.0
bigscience-openrail-mWizardCoder: Empowering Code Large Language Models with Evol-Instruct π Home Page π€ HF Repo β’π± Github Repo β’ π¦ Twitter π WizardLM β’ π ...
PygmalionAI/pygmalion-6b
creativeml-openrail-mPygmalion 6B Model description Pymalion 6B is a proof-of-concept dialogue model based on EleutherAI's GPT-J-6B( Warning: This model is NOT s...
google/gemma-2b-it
gemmaVisit HuggingFace for more details.
Gustavosta/MagicPrompt-Stable-Diffusion
mitMagicPrompt - Stable Diffusion This is a model from the MagicPrompt series of models, which are GPT-2( models intended to generate prompt te...
deepseek-ai/DeepSeek-R1-Distill-Llama-8B
mitDeepSeek-R1 Paper LinkποΈ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
mistralai/Mixtral-8x22B-Instruct-v0.1
apache-2.0Model Card for Mixtral-8x22B-Instruct-v0.1 Encode and Decode with mistralcommon py from mistralcommon.tokens.tokenizers.mistral import Mistr...