LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
This repository is primarily for the Transformers framework. If you're using other open-source frameworks, please use the alternative repository: MiniMax-M1-80k
Step 3.5 Flash (visit website) is our most capable open-source foundation model, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency. We also open-sourced the training codebase (SteptronOss), with support for continue pretrain, SFT, RL (W…
🤗 Hugging Face 🤖 ModelScope 🐙 Experience Now
🤗 Hugging Face      🤖 ModelScope 🐙 Experience Now
💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat
🤗 HuggingFace 🤖 ModelScope 🪡 AngelSlim
Ling is a MoE LLM provided and open-sourced by InclusionAI. We introduce two different sizes, which are Ling-Lite and Ling-Plus. Ling-Lite has 16.8 billion parameters with 2.75 billion activated parameters, while Ling-Plus has 290 billion parameters with 28.8 billion activated pa…
🤗 Hugging Face 🖥️ Official Website 🕖 HunyuanAPI 🕹️ Demo 🤖 ModelScope
This repository is primarily for the Transformers framework. If you're using other open-source frameworks, please use the alternative repository: MiniMax-M1-40k
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]
[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.
[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.