google/gemma-3-4b-it-qat-q4_0-gguf

gemma

Visit HuggingFace for more details.

image text to textBy google

microsoft/Florence-2-large

mit

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks Model Summary This is a continued pretrained version of Florenc...

image text to textBy microsoft

stepfun-ai/GOT-OCR2_0

apache-2.0

General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model 🔋Online Demo( | 🌟GitHub( | 📜Paper( Haoran Wei( Chenglong Liu, Jinyue C...

image text to textBy stepfun-ai

meta-llama/Llama-3.2-11B-Vision-Instruct

llama3.2

Visit HuggingFace for more details.

image text to textBy meta-llama

ds4sd/SmolDocling-256M-preview

cdla-permissive-2.0

📢 New Release: We’ve released granite-docling-258M, the successor to SmolDocling. It will now receive updates and support, check it out! Sm...

image text to textBy ds4sd

mistralai/Mistral-Small-3.1-24B-Instruct-2503

apache-2.0

Model Card for Mistral-Small-3.1-24B-Instruct-2503 Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art visi...

image text to textBy mistralai

Qwen/Qwen2-VL-7B-Instruct

apache-2.0

Qwen2-VL-7B-Instruct Introduction We're excited to unveil Qwen2-VL, the latest iteration of our Qwen-VL model, representing nearly a year of...

image text to textBy Qwen

vikhyatk/moondream2

apache-2.0

⚠️ This repository contains the latest version of Moondream 2, our previous generation model. The latest version of Moondream is Moondream 3...

image text to textBy vikhyatk

openbmb/MiniCPM-V-2_6

Visit HuggingFace for more details.

image text to textBy openbmb

meta-llama/Llama-4-Scout-17B-16E-Instruct

other

Visit HuggingFace for more details.

image text to textBy meta-llama
Showing 10 of 159