Sao10K/L3-8B-Stheno-v3.2
cc-by-nc-4.0Just message me on discord if you want to host this privately for a service or something. We can talk. Train used 1x H100 SXM for like a tot...
nvidia/Llama-3_3-Nemotron-Super-49B-v1
otherLlama-3.3-Nemotron-Super-49B-v1 Model Overview !Accuracy Comparison Plot(./accuracyplot.png) Llama-3.3-Nemotron-Super-49B-v1 is a large lang...
LGAI-EXAONE/EXAONE-Deep-32B
otherEXAONE-Deep-32B Introduction We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and co...
Qwen/Qwen3-0.6B
apache-2.0Qwen3-0.6B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense ...
epfl-llm/meditron-7b
llama2Visit HuggingFace for more details.
microsoft/DialoGPT-large
mitA State-of-the-Art Large-scale Pretrained Response generation model (DialoGPT) DialoGPT is a SOTA large-scale pretrained dialogue response g...
OpenAssistant/oasst-sft-1-pythia-12b
apache-2.0Open-Assistant SFT-1 12B Model This is the first iteration English supervised-fine-tuning (SFT) model of the Open-Assistant( project. It is ...
google/gemma-1.1-7b-it
gemmaVisit HuggingFace for more details.
Qwen/Qwen2.5-0.5B
apache-2.0Qwen2.5-0.5B Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language mode...
unsloth/DeepSeek-R1-Distill-Llama-8B-GGUF
llama3.1See our collection for versions of Deepseek-R1 including GGUF & 4-bit formats. Unsloth's DeepSeek-R1 1.58-bit + 2-bit Dynamic Quants is sele...