Byte Goose AIBeta
Xkev/Llama-3.2V-11B-cot
apache-2.0Model Card for Model ID Llama-3.2V-11B-cot is a visual language model capable of spontaneous, systematic reasoning. The model was proposed i...
image text to textBy Xkev
gemma-4-26B-A4B-it
apache-2.0Official image-text-to-text model by google.
image text to textBy google
google/gemma-3-27b-it
gemmaVisit HuggingFace for more details.
image text to textBy google
microsoft/OmniParser
mit📢 Project Page( Blog Post( Demo( Model Summary OmniParser is a general screen parsing tool, which interprets/converts UI screenshot to stru...
image text to textBy microsoft
google/gemma-3-27b-it
gemmaVisit HuggingFace for more details.
image text to textBy google
gemma-4-31B-it
apache-2.0Official image-text-to-text model by google.
image text to textBy google
Qwen2.5-VL-7B-Instruct
apache-2.0Official image-text-to-text model by Qwen.
image text to textBy Qwen
Qwen3.5-9B
apache-2.0Official image-text-to-text model by Qwen.
image text to textBy Qwen
Qwen3.6-35B-A3B-FP8
apache-2.0Official image-text-to-text model by Qwen.
image text to textBy Qwen
Qwen3.5-4B
apache-2.0Official image-text-to-text model by Qwen.
image text to textBy Qwen