Sao10K/L3-8B-Stheno-v3.2

cc-by-nc-4.0

Just message me on discord if you want to host this privately for a service or something. We can talk. Train used 1x H100 SXM for like a tot...

text generationBy Sao10K

nvidia/Llama-3_3-Nemotron-Super-49B-v1

other

Llama-3.3-Nemotron-Super-49B-v1 Model Overview !Accuracy Comparison Plot(./accuracyplot.png) Llama-3.3-Nemotron-Super-49B-v1 is a large lang...

text generationBy nvidia

LGAI-EXAONE/EXAONE-Deep-32B

other

EXAONE-Deep-32B Introduction We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and co...

text generationBy LGAI-EXAONE

Qwen/Qwen3-0.6B

apache-2.0

Qwen3-0.6B Qwen3 Highlights Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense ...

text generationBy Qwen

epfl-llm/meditron-7b

llama2

Visit HuggingFace for more details.

text generationBy epfl-llm

microsoft/DialoGPT-large

mit

A State-of-the-Art Large-scale Pretrained Response generation model (DialoGPT) DialoGPT is a SOTA large-scale pretrained dialogue response g...

text generationBy microsoft

OpenAssistant/oasst-sft-1-pythia-12b

apache-2.0

Open-Assistant SFT-1 12B Model This is the first iteration English supervised-fine-tuning (SFT) model of the Open-Assistant( project. It is ...

text generationBy OpenAssistant

google/gemma-1.1-7b-it

gemma

Visit HuggingFace for more details.

text generationBy google

Qwen/Qwen2.5-0.5B

apache-2.0

Qwen2.5-0.5B Introduction Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language mode...

text generationBy Qwen

unsloth/DeepSeek-R1-Distill-Llama-8B-GGUF

llama3.1

See our collection for versions of Deepseek-R1 including GGUF & 4-bit formats. Unsloth's DeepSeek-R1 1.58-bit + 2-bit Dynamic Quants is sele...

text generationBy unsloth
Showing 10 of 563