AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

377 models for "MoE" Compare
🤖
MiniMax-M1-80k-hf
MiniMaxAI

This repository is primarily for the Transformers framework. If you're using other open-source frameworks, please use the alternative repository: MiniMax-M1-80k

Open Source ↓ 329
🤖
Step-3.5-Flash-Base
stepfun-ai

Step 3.5 Flash (visit website) is our most capable open-source foundation model, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency. We also open-sourced the training codebase (SteptronOss), with support for continue pretrain, SFT, RL (W…

Open Source ↓ 321
🤖
Ring-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 Experience Now

Open Source ↓ 313
🤖
Ring-mini-2.0
inclusionAI

🤗 Hugging Face &nbsp&nbsp &nbsp&nbsp🤖 ModelScope   🐙 Experience Now

Open Source ↓ 306
🤖
Intern-S1-FP8
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat

Open Source ↓ 304
🤖
Hunyuan-0.5B-Pretrain
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 0.5B ↓ 288
🤖
Ling-plus
inclusionAI

Ling is a MoE LLM provided and open-sourced by InclusionAI. We introduce two different sizes, which are Ling-Lite and Ling-Plus. Ling-Lite has 16.8 billion parameters with 2.75 billion activated parameters, while Ling-Plus has 290 billion parameters with 28.8 billion activated pa…

Open Source ↓ 265
🤖
Hunyuan-A13B-Instruct-FP8
tencent

🤗  Hugging Face       🖥️  Official Website       🕖  HunyuanAPI       🕹️  Demo       🤖  ModelScope

Open Source 13.0B ↓ 259
🤖
MiniMax-M1-40k-hf
MiniMaxAI

This repository is primarily for the Transformers framework. If you're using other open-source frameworks, please use the alternative repository: MiniMax-M1-40k

Open Source ↓ 253
🤖
AquilaMoE
BAAI

AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]

Open Source ↓ 241
🤖
ERNIE-4.5-VL-424B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 239
🤖
ERNIE-4.5-0.3B-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 0.3B ↓ 228