nvidia/segformer-b0-finetuned-ade-512-512
otherSegFormer (b0-sized) model fine-tuned on ADE20k SegFormer model fine-tuned on ADE20k at resolution 512x512. It was introduced in the paper S...
TencentBAC/Conan-embedding-v1
cc-by-nc-4.0Conan-embedding-v1 Performance | Model | Average | CLS | Clustering | Reranking | Retrieval | STS | PairCLS | | :-------------------: | :---...
LGAI-EXAONE/EXAONE-3.5-2.4B-Instruct
otherEXAONE-3.5-2.4B-Instruct Introduction We introduce EXAONE 3.5, a collection of instruction-tuned bilingual (English and Korean) generative m...
openlm-research/open_llama_3b
apache-2.0OpenLLaMA: An Open Reproduction of LLaMA In this repo, we present a permissively licensed open source reproduction of Meta AI's LLaMA( large...
Zhengyi/LLaMA-Mesh
llama3.1LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models Paper( | Project Page( Pre-trained model weights of LLaMA-Mesh: Unifying 3D Mes...
BlinkDL/rwkv-4-pile-7b
apache-2.0RWKV-4 7B UPDATE: Try RWKV-4-World ( for generation & chat & code in 100+ world languages, with great English zero-shot & in-context learnin...
Salesforce/blip-vqa-base
bsd-3-clauseBLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for BLIP trained on visu...
tianweiy/DMD2
cc-by-nc-4.0DMD2 Model Card !image/jpeg( > Improved Distribution Matching Distillation for Fast Image Synthesis( > Tianwei Yin, Michaël Gharbi, Taesung ...
deepseek-ai/deepseek-vl2-small
other1. Introduction Introducing DeepSeek-VL2, an advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly imp...
nvidia/Mistral-NeMo-12B-Instruct
apache-2.0Mistral-NeMo-12B-Instruct !Model architecture( size( Model Overview: Mistral-NeMo-12B-Instruct is a Large Language Model (LLM) composed of 1...