Qwen/CodeQwen1.5-7B-Chat
otherCodeQwen1.5-7B-Chat Introduction CodeQwen1.5 is the Code-Specific version of Qwen1.5. It is a transformer-based decoder-only language model ...
Qwen/Qwen3-32B
apache-2.0Qwen3-32B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense a...
nsfw_image_detection
apache-2.0Official image-classification model by Falconsai.
MLP-KTLim/llama-3-Korean-Bllossom-8B
llama3Update! ~~2024.08.09 Llama3.1 버전을 기반으로한 Bllossom-8B로 모델을 업데이트 했습니다. 기존 llama3기반 Bllossom 보다 평균 5%정도 성능 향상이 있었습니다.~~(수정중에 있습니다.) 2024.06.18 사...
apple/OpenELM-3B-Instruct
apple-amlrOpenELM Sachin Mehta, Mohammad Hossein Sekhavat, Qingqing Cao, Maxwell Horton, Yanzi Jin, Chenfan Sun, Iman Mirzadeh, Mahyar Najibi, Dmitry ...
bigcode/santacoder
bigcode-openrail-mSantaCoder !banner( Play with the model on the SantaCoder Space Demo( Table of Contents 1. Model Summary(model-summary) 2. Use(use) 3. Limit...
google/vit-base-patch16-224-in21k
apache-2.0Vision Transformer (base-sized model) Vision Transformer (ViT) model pre-trained on ImageNet-21k (14 million images, 21,843 classes) at reso...
deepseek-ai/deepseek-vl2
other1. Introduction Introducing DeepSeek-VL2, an advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly imp...
Wan-AI/Wan2.1-T2V-1.3B
apache-2.0Wan2.1 💜 Wan    |    🖥️ GitHub    |   🤗 Hugging Face   |   🤖 ModelScope   | &nbs...
hfl/chinese-roberta-wwm-ext
apache-2.0Please use 'Bert' related functions to load this model! Chinese BERT with Whole Word Masking For further accelerating Chinese natural langua...