THUDM/GLM-4-32B-0414
mitVisit HuggingFace for more details.
LGAI-EXAONE/EXAONE-3.0-7.8B-Instruct
otherVisit HuggingFace for more details.
ai-forever/ruGPT-3.5-13B
mit🗿 ruGPT-3.5 13B Language model for Russian. Model has 13B parameters as you can guess from it's name. This is our biggest model so far and ...
PowerInfer/SmallThinker-3B-Preview
SmallThinker-3B-preview We introduce SmallThinker-3B-preview, a new model fine-tuned from the Qwen2.5-3b-Instruct( model. Now you can direct...
microsoft/DialoGPT-medium
mitA State-of-the-Art Large-scale Pretrained Response generation model (DialoGPT) DialoGPT is a SOTA large-scale pretrained dialogue response g...
all-hands/openhands-lm-32b-v0.1
mitOpenHands LM v0.1 Blog • Use it in OpenHands --- Autonomous agents for software development are already contributing to a wide range of soft...
01-ai/Yi-6B
apache-2.0Building the Next Generation of Open-Source and Bilingual LLMs 🤗 Hugging Face • 🤖 ModelScope • ✡️ WiseModel 👩🚀 Ask questions or discuss...
meta-llama/Llama-3.1-70B
llama3.1Visit HuggingFace for more details.
defog/sqlcoder-7b-2
cc-by-sa-4.0Update notice The model weights were updated at 7 AM UTC on Feb 7, 2024. The new model weights lead to a much more performant model – partic...
JetBrains/Mellum-4b-base
apache-2.0Model Description Mellum-4b-base is JetBrains' first open-source large language model (LLM) optimized for code-related tasks. Trained on ove...