bosonai/Higgs-Llama-3-70B
otherHiggs-Llama-3-70B Higgs-Llama-3-70B is post-trained from meta-llama/Meta-Llama-3-70B( specially tuned for role-playing while being competiti...
microsoft/Orca-2-7b
otherOrca 2 Orca 2 is built for research purposes only and provides a single turn response in tasks such as reasoning over user given data, readi...
huihui-ai/DeepSeek-R1-Distill-Qwen-32B-abliterated
huihui-ai/DeepSeek-R1-Distill-Qwen-32B-abliterated This is an uncensored version of deepseek-ai/DeepSeek-R1-Distill-Qwen-32B( created with a...
McGill-NLP/Llama-3-8B-Web
llama3Llama-3-8B-Web 💻 GitHub 🏠 Homepage 🤗 Llama-3-8B-Web By using this model, you are accepting the terms of the Meta Llama 3 Community Licens...
uer/gpt2-chinese-cluecorpussmall
Chinese GPT2 Models Model description The set of GPT2 models, except for GPT2-xlarge model, are pre-trained by UER-py( which is introduced i...
MarinaraSpaghetti/NemoMix-Unleashed-12B
apache-2.0!image/jpeg( !image/png( Information Details Okay, I tried really hard to improve my ChatML merges, but that has gone terribly wrong. Everyo...
arcee-ai/SuperNova-Medius
apache-2.0Arcee-SuperNova-Medius Arcee-SuperNova-Medius is a 14B parameter language model developed by Arcee.ai, built on the Qwen2.5-14B-Instruct arc...
Qwen/Qwen3-4B
apache-2.0Qwen3-4B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense an...
EleutherAI/gpt-neo-125m
mitGPT-Neo 125M Model Description GPT-Neo 125M is a transformer model designed using EleutherAI's replication of the GPT-3 architecture. GPT-Ne...
aaditya/Llama3-OpenBioLLM-8B
llama3!image/png( Advancing Open-source Large Language Models in Medical Domain Online Demo | GitHub | Paper | Discord !image/jpeg( Introducing Op...