AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

375 models for "MoE" Compare
🤖
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4
nvidia

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4

Open Source 30.0B ★ 24.0 ↓ 946.3K
🤖
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
nvidia

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

Open Source 30.0B ★ 24.0 ↓ 438.3K
🤖
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark
nvidia

The NVIDIA Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark model is the DSpark speculative decoding checkpoint for NVIDIA's Nemotron-3.5-Lightning-30B-A3B model family, which is a hybrid LatentMoE language model designed for reasoning, chat, and agentic workflows. For more informatio…

Open Source 30.0B ★ 24.0 ↓ 200K
🤖
OpenAI: gpt-oss-120b (batch)
openai

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Closed Source ★ 24.0
🤖
K-EXAONE-236B-A23B
LGAI-EXAONE

Check out the NVFP4+GPTQ weights by FuriosaAI ! ➡️ link

Open Source 236.0B ★ 22.0 ↓ 21K
🤖
K-EXAONE-236B-A23B-FP8
LGAI-EXAONE

Check out the NVFP4+GPTQ weights by FuriosaAI ! ➡️ link

Open Source 236.0B ★ 22.0 ↓ 2.7K
🤖
Qwen3-Coder-Next-FP8
Qwen

Today, we're announcing Qwen3-Coder-Next-FP8 , an open-weight language model designed specifically for coding agents and local development. It features the following key enhancements:

Code 80.0B ★ 21.0 ↓ 1.8M
🤖
EXAONE-4.5-33B
LGAI-EXAONE

We introduce EXAONE 4.5, the first open-weight vision language model developed by LG AI Research. Integrating a dedicated visual encoder into the existing EXAONE 4.0 framework, we expand the model's capability toward multimodality. EXAONE 4.5 features 33 billion parameters in tot…

Open Source 33.0B ★ 21.0 ↓ 79.8K
🤖
EXAONE-4.5-33B-AWQ
LGAI-EXAONE

We introduce EXAONE 4.5, the first open-weight vision language model developed by LG AI Research. Integrating a dedicated visual encoder into the existing EXAONE 4.0 framework, we expand the model's capability toward multimodality. EXAONE 4.5 features 33 billion parameters in tot…

Multimodal 33.0B ★ 21.0 ↓ 52.1K
🤖
EXAONE-4.5-33B-FP8
LGAI-EXAONE

We introduce EXAONE 4.5, the first open-weight vision language model developed by LG AI Research. Integrating a dedicated visual encoder into the existing EXAONE 4.0 framework, we expand the model's capability toward multimodality. EXAONE 4.5 features 33 billion parameters in tot…

Open Source 33.0B ★ 21.0 ↓ 3K
🤖
Qwen: Qwen3 Coder 480B A35B
qwen

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

Code ★ 21.0
🤖
Qwen: Qwen3 Coder Next
qwen

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

Code ★ 21.0