nvidia/Nemotron-Mini-4B-Instruct
otherNemotron-Mini-4B-Instruct Model Overview Nemotron-Mini-4B-Instruct is a model for generating responses for roleplaying, retrieval augmented ...
LGAI-EXAONE/EXAONE-3.5-2.4B-Instruct
otherEXAONE-3.5-2.4B-Instruct Introduction We introduce EXAONE 3.5, a collection of instruction-tuned bilingual (English and Korean) generative m...
openlm-research/open_llama_3b
apache-2.0OpenLLaMA: An Open Reproduction of LLaMA In this repo, we present a permissively licensed open source reproduction of Meta AI's LLaMA( large...
BlinkDL/rwkv-4-pile-7b
apache-2.0RWKV-4 7B UPDATE: Try RWKV-4-World ( for generation & chat & code in 100+ world languages, with great English zero-shot & in-context learnin...
Groq/Llama-3-Groq-70B-Tool-Use
llama3Llama-3-70B-Tool-Use This is the 70B parameter version of the Llama 3 Groq Tool Use model, specifically designed for advanced tool use and f...
qnguyen3/nanoLLaVA
apache-2.0nanoLLaVA - Sub 1B Vision-Language Model IMPORTANT: nanoLLaVA-1.5 is out with a much better performance. Please find it here( Description na...
openchat/openchat-3.6-8b-20240522
llama3Advancing Open-source Language Models with Mixed-Quality Data Online Demo | GitHub | Paper | Discord Sponsored by RunPod Llama 3 Version: OP...
facebook/galactica-120b
cc-by-nc-4.0!logo( GALACTICA 120 B (huge) Model card from the original repo( Following Mitchell et al. (2018)( this model card provides information abou...
Qwen/Qwen2.5-32B
apache-2.0Qwen2.5-32B Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language model...
zl111/ChatDoctor
gplVisit HuggingFace for more details.