mistralai/Mistral-Small-3.1-24B-Instruct-2503

apache-2.0

Model Card for Mistral-Small-3.1-24B-Instruct-2503 Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art visi...

image text to textBy mistralai

Qwen/Qwen2-VL-7B-Instruct

apache-2.0

Qwen2-VL-7B-Instruct Introduction We're excited to unveil Qwen2-VL, the latest iteration of our Qwen-VL model, representing nearly a year of...

image text to textBy Qwen

vikhyatk/moondream2

apache-2.0

⚠️ This repository contains the latest version of Moondream 2, our previous generation model. The latest version of Moondream is Moondream 3...

image text to textBy vikhyatk

openbmb/MiniCPM-V-2_6

Visit HuggingFace for more details.

image text to textBy openbmb

google/gemma-3-27b-it

gemma

Visit HuggingFace for more details.

image text to textBy google

google/gemma-3-4b-it

gemma

Visit HuggingFace for more details.

image text to textBy google

reducto/RolmOCR

apache-2.0

RolmOCR by Reducto AI( Earlier this year, the Allen Institute for AI( released olmOCR, an open-source tool that performs document OCR using ...

image text to textBy reducto

HuggingFaceTB/SmolVLM2-2.2B-Instruct

apache-2.0

SmolVLM2 2.2B SmolVLM2-2.2B is a lightweight multimodal model designed to analyze video content. The model processes videos, images, and tex...

image text to textBy HuggingFaceTB

HuggingFaceTB/SmolVLM-500M-Instruct

apache-2.0

SmolVLM-500M SmolVLM-500M is a tiny multimodal model, member of the SmolVLM family. It accepts arbitrary sequences of image and text inputs ...

image text to textBy HuggingFaceTB

allenai/MolmoE-1B-0924

apache-2.0

MolmoE 1B Molmo is a family of open vision-language models developed by the Allen Institute for AI. Molmo models are trained on PixMo, a dat...

image text to textBy allenai
Showing 10 of 159