AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

377 models for "MoE" Compare
🤖
North-Mini-Code-1.0
CohereLabs

North Mini Code is an open weights research release of a 30B-A3B parameter model optimized for code generation, agentic software engineering, and terminal tasks.

Code ★ 20.0 ↓ 21.5K
🤖
North-Mini-Code-1.0-fp8
CohereLabs

North Mini Code is an open weights research release of a 30B-A3B parameter model optimized for code generation, agentic software engineering, and terminal tasks.

Code ★ 20.0 ↓ 1.5K
🤖
North-Mini-Code-1.0-w4a16
CohereLabs

North Mini Code is an open weights research release of a 30B-A3B parameter model optimized for code generation, agentic software engineering, and terminal tasks.

Code ★ 20.0 ↓ 1.1K
🤖
Kimi-K2-Thinking
moonshotai

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, we built it as a thinking agent that reasons step-by-step while dynamically invoking tools. It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and oth…

Open Source ★ 17.0 ↓ 43.1K
🤖
LongCat-Flash-Lite
meituan-longcat

Model Introduction We introduce LongCat-Flash-Lite, a non-thinking 68.5B parameter Mixture-of-Experts (MoE) model with approximately 3B activated parameters, supporting a 256k context length through the YaRN method. Building upon the LongCat-Flash architecture, LongCat-Flash-Lite…

Open Source ★ 17.0 ↓ 15.4K
🤖
LongCat-Flash-Lite-FP8
meituan-longcat

Model Introduction We introduce LongCat-Flash-Lite, a non-thinking 68.5B parameter Mixture-of-Experts (MoE) model with approximately 3B activated parameters, supporting a 256k context length through the YaRN method. Building upon the LongCat-Flash architecture, LongCat-Flash-Lite…

Open Source 68.5B ★ 17.0 ↓ 14.3K
🤖
LongCat-Flash-Lite-Sparse
meituan-longcat

LongCat-Flash-Lite-Sparse is a non-thinking Mixture-of-Experts (MoE) model with 69B total parameters and approximately 3B activated parameters per token. Built on LongCat-Flash-Lite, it replaces dense MLA with LongCat Sparse Attention (LSA) and natively supports context lengths o…

Open Source ★ 17.0 ↓ 1.3K
🤖
gpt-oss-20b
openai

Try gpt-oss · Guides · Model card · OpenAI blog

Open Source 20.0B ★ 15.0 ↓ 6.5M
🤖
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 15.0 ↓ 871K
🤖
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 15.0 ↓ 685.6K
🤖
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 15.0 ↓ 580.3K
🤖
NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16
nvidia

NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16

Open Source 30.0B ★ 15.0 ↓ 121.6K