AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
GLM-5.2-NVFP4 logo
GLM-5.2-NVFP4
nvidia

Description: The NVIDIA GLM-5.2 NVFP4 model is the quantized version of ZAI’s GLM-5.2 model, which is an auto-regressive language model that uses an optimized transformer architecture. GLM-5.2 is a Mixture-of-Experts (MoE) model for reasoning and coding that uses sparse attention…

Open Source ★ 53.0 ↓ 450.6K
DeepSeek-V4-Pro-NVFP4 logo
DeepSeek-V4-Pro-NVFP4
nvidia

Description: The NVIDIA DeepSeek-V4-Pro-NVFP4 model is the quantized version of the DeepSeek-V4-Pro model, which is a Mixture-of-Experts (MoE) language model with 1.6 trillion total parameters and 49 billion activated parameters. For more information, please check here. The NVIDI…

Open Source ★ 53.0 ↓ 146.5K
DeepSeek-V4-Flash-DSpark logo
DeepSeek-V4-Flash-DSpark
deepseek-ai

DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence

Open Source ★ 52.0 ↓ 579.7K
DeepSeek-V4-Flash-NVFP4 logo
DeepSeek-V4-Flash-NVFP4
nvidia

Description: The NVIDIA DeepSeek-V4-Flash-NVFP4 model is a quantized version of DeepSeek AI's DeepSeek-V4-Flash model, an autoregressive Mixture-of-Experts language model that uses an optimized Transformer architecture with hybrid attention (Compressed Sparse Attention and Heavil…

Open Source ★ 52.0 ↓ 223.1K
DeepSeek: DeepSeek V4 Flash Vision Exp logo
DeepSeek: DeepSeek V4 Flash Vision Exp
deepseek

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

Multimodal ★ 51.0
GLM-5.3 logo
GLM-5.3
zai-org

GLM-5.3 uses the same base model as GLM-5.2 — every gain comes from post-training. Compared with GLM-5.2, it is much better at complex coding and long-horizon tasks:

Open Source ★ 45.0 ↓ 1.5M
MiniMax-M3-NVFP4 logo
MiniMax-M3-NVFP4
nvidia

Description MiniMax-M3 is a multimodal model with frontier-level coding and agentic capabilities, built on a Mixture-of-Experts architecture with a 1M-token context window. The model processes text, image, video, and computer use inputs and produces text outputs, with emphasis on…

Open Source ★ 45.0 ↓ 183K
GLM-5.3-BF16 logo
GLM-5.3-BF16
zai-org

GLM-5.3 uses the same base model as GLM-5.2 — every gain comes from post-training. Compared with GLM-5.2, it is much better at complex coding and long-horizon tasks:

Open Source ★ 45.0 ↓ 38.9K
Kimi-K2.7-Code-NVFP4 logo
Kimi-K2.7-Code-NVFP4
nvidia

Description: The NVIDIA Kimi-K2.7-Code NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2.7-Code model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2.7-Code NV…

Code ★ 43.0 ↓ 294.8K
MiMo-V2.5-Pro logo
MiMo-V2.5-Pro
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio  

Open Source ★ 43.0 ↓ 21.9K
MiMo-V2.5-Pro-Base logo
MiMo-V2.5-Pro-Base
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio  

Open Source ★ 43.0 ↓ 835
GLM-5.3-Flash logo
GLM-5.3-Flash
zai-org

👋 Join our WeChat or Discord community. 📖 Check out the GLM-5.3-Flash blog and GLM-5 Technical report . 📍 Use GLM-5.3-Flash API services on Z.ai API Platform.

Open Source ★ 42.0 ↓ 6.4M