AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

375 models for "MoE" Compare
🤖
Step-3.7-Flash-FP8
stepfun-ai

[ModelPage] : https://static.stepfun.com/blog/step-3.7-flash/

Open Source ★ 31.0 ↓ 8.7K
🤖
Step-3.7-Flash-NVFP4
stepfun-ai

[ModelPage] : https://static.stepfun.com/blog/step-3.7-flash/

Open Source ★ 31.0 ↓ 5.4K
🤖
K-EXAONE-2.0-750B-A37B
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 4.1K
🤖
K-EXAONE-2.0-750B-A37B-NVFP4
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 1.9K
🤖
K-EXAONE-2.0-750B-A37B-FP8
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 1.9K
🤖
K-EXAONE-2.0-750B-A37B-DSpark
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 1.1K
🤖
gemma-4-31B-it
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 30.0 ↓ 8.3M
🤖
gemma-4-31B-it-qat-w4a16-ct
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 30.0 ↓ 1.2M
🤖
gemma-4-31B
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 30.0 ↓ 717.3K
🤖
gemma-4-26B-A4B-it
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 26.0B ★ 26.0 ↓ 8.2M
🤖
Gemma-4-26B-A4B-NVFP4
nvidia

Description: Gemma 4 26B IT is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output. It is designed to deliver frontier-level performance for reasoning, agentic workflows, coding,…

Open Source 26.0B ★ 26.0 ↓ 1.5M
🤖
NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
nvidia

:--- :--- Total Parameters 120B (12B active) Architecture LatentMoE - Mamba-2 + MoE + Attention hybrid with Multi-Token Prediction (MTP) Context Length Up to 1M tokens Minimum GPU Requirement 1× B200 OR 1× DGX Spark Supported Languages English, French, German, Italian, Japanese,…

Reasoning 120.0B ★ 26.0 ↓ 1.2M