nvidia/segformer-b0-finetuned-ade-512-512

other

SegFormer (b0-sized) model fine-tuned on ADE20k SegFormer model fine-tuned on ADE20k at resolution 512x512. It was introduced in the paper S...

image segmentationBy nvidia

TencentBAC/Conan-embedding-v1

cc-by-nc-4.0

Conan-embedding-v1 Performance | Model | Average | CLS | Clustering | Reranking | Retrieval | STS | PairCLS | | :-------------------: | :---...

GeneralBy TencentBAC

LGAI-EXAONE/EXAONE-3.5-2.4B-Instruct

other

EXAONE-3.5-2.4B-Instruct Introduction We introduce EXAONE 3.5, a collection of instruction-tuned bilingual (English and Korean) generative m...

text generationBy LGAI-EXAONE

openlm-research/open_llama_3b

apache-2.0

OpenLLaMA: An Open Reproduction of LLaMA In this repo, we present a permissively licensed open source reproduction of Meta AI's LLaMA( large...

text generationBy openlm-research

Zhengyi/LLaMA-Mesh

llama3.1

LLaMA-Mesh: Unifying 3D Mesh Generation with Language Models Paper( | Project Page( Pre-trained model weights of LLaMA-Mesh: Unifying 3D Mes...

text to 3dBy Zhengyi

BlinkDL/rwkv-4-pile-7b

apache-2.0

RWKV-4 7B UPDATE: Try RWKV-4-World ( for generation & chat & code in 100+ world languages, with great English zero-shot & in-context learnin...

text generationBy BlinkDL

Salesforce/blip-vqa-base

bsd-3-clause

BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for BLIP trained on visu...

visual question answeringBy Salesforce

tianweiy/DMD2

cc-by-nc-4.0

DMD2 Model Card !image/jpeg( > Improved Distribution Matching Distillation for Fast Image Synthesis( > Tianwei Yin, Michaël Gharbi, Taesung ...

text to imageBy tianweiy

deepseek-ai/deepseek-vl2-small

other

1. Introduction Introducing DeepSeek-VL2, an advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly imp...

image text to textBy deepseek-ai

nvidia/Mistral-NeMo-12B-Instruct

apache-2.0

Mistral-NeMo-12B-Instruct !Model architecture( size( Model Overview: Mistral-NeMo-12B-Instruct is a Large Language Model (LLM) composed of 1...

GeneralBy nvidia
Showing 10 of 1932