mistralai/Mistral-Small-3.1-24B-Instruct-2503
apache-2.0Model Card for Mistral-Small-3.1-24B-Instruct-2503 Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art visi...
Qwen/Qwen2-VL-7B-Instruct
apache-2.0Qwen2-VL-7B-Instruct Introduction We're excited to unveil Qwen2-VL, the latest iteration of our Qwen-VL model, representing nearly a year of...
vikhyatk/moondream2
apache-2.0⚠️ This repository contains the latest version of Moondream 2, our previous generation model. The latest version of Moondream is Moondream 3...
openbmb/MiniCPM-V-2_6
Visit HuggingFace for more details.
google/gemma-3-27b-it
gemmaVisit HuggingFace for more details.
google/gemma-3-4b-it
gemmaVisit HuggingFace for more details.
reducto/RolmOCR
apache-2.0RolmOCR by Reducto AI( Earlier this year, the Allen Institute for AI( released olmOCR, an open-source tool that performs document OCR using ...
HuggingFaceTB/SmolVLM2-2.2B-Instruct
apache-2.0SmolVLM2 2.2B SmolVLM2-2.2B is a lightweight multimodal model designed to analyze video content. The model processes videos, images, and tex...
HuggingFaceTB/SmolVLM-500M-Instruct
apache-2.0SmolVLM-500M SmolVLM-500M is a tiny multimodal model, member of the SmolVLM family. It accepts arbitrary sequences of image and text inputs ...
allenai/MolmoE-1B-0924
apache-2.0MolmoE 1B Molmo is a family of open vision-language models developed by the Allen Institute for AI. Molmo models are trained on PixMo, a dat...