jinaai/ReaderLM-v2
cc-by-nc-4.0Trained by Jina AI. Blog( | API( | Colab( | AWS( | Azure( Arxiv( ReaderLM-v2 ReaderLM-v2 is a 1.5B parameter language model that converts ra...
google/gemma-2-9b
gemmaVisit HuggingFace for more details.
deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
mitDeepSeek-R1 Paper Link👁️ 1. Introduction We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-...
HuggingFaceH4/zephyr-7b-beta
mitModel Card for Zephyr 7B β Zephyr is a series of language models that are trained to act as helpful assistants. Zephyr-7B-β is the second mo...
agentica-org/DeepCoder-14B-Preview
mitDeepCoder-14B-Preview 🚀 Democratizing Reinforcement Learning for LLMs (RLLM) 🌟 DeepCoder Overview DeepCoder-14B-Preview is a code reasonin...
allenai/OLMo-7B
apache-2.0!mof-class1-qualified( Model Card for OLMo 7B For transformers versions v4.40.0 or newer, we suggest using OLMo 7B HF( instead. OLMo is a se...
deepseek-ai/DeepSeek-Coder-V2-Instruct
otherAPI Platform | How to Use | License | Paper Link👁️ DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence 1. ...
upstage/SOLAR-10.7B-Instruct-v1.0
cc-by-nc-4.0Meet 10.7B Solar: Elevating Performance with Upstage Depth UP Scaling! (This model is upstage/SOLAR-10.7B-v1.0( fine-tuned version for singl...
HuggingFaceTB/SmolLM2-1.7B-Instruct
apache-2.0SmolLM2 !image/png( Table of Contents 1. Model Summary(model-summary) 2. Evaluation(evaluation) 3. Examples(examples) 4. Limitations(limitat...
bigcode/starcoder2-15b
bigcode-openrail-mStarCoder2 Table of Contents 1. Model Summary(model-summary) 2. Use(use) 3. Limitations(limitations) 4. Training(training) 5. License(licens...