AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 logo
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16
nvidia

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

Open Source 30.0B ★ 13.0 ↓ 599K
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark logo
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark
nvidia

The NVIDIA Nemotron-3.5-Lightning-30B-A3B-NVFP4-DSpark model is the DSpark speculative decoding checkpoint for NVIDIA's Nemotron-3.5-Lightning-30B-A3B model family, which is a hybrid LatentMoE language model designed for reasoning, chat, and agentic workflows. For more informatio…

Open Source 30.0B ★ 13.0 ↓ 134.1K
Kimi-K2-Thinking-NVFP4 logo
Kimi-K2-Thinking-NVFP4
nvidia

Description: The NVIDIA Kimi-K2-Thinking-NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2-Thinking model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2-Think…

Reasoning 1000.0B ★ 11.0 ↓ 166K
granite-4.2-8b logo
granite-4.2-8b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-8B-Base Parameters 8B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Engli…

Open Source 8.0B ★ 11.0 ↓ 133.6K
diffusiongemma-26B-A4B-it-NVFP4 logo
diffusiongemma-26B-A4B-it-NVFP4
nvidia

--- pipeline tag: text-generation base model: google/diffusiongemma-26B-A4B-it license: apache-2.0 license name: apache-license-2.0 license link: https://ai.google.dev/gemma/apache 2 tags: - nvidia - ModelOpt - DiffusionGemma-26B-A4B-IT - quantized - NVFP4 - nvfp4 ---

Open Source 26.0B ★ 10.0 ↓ 92.2K
granite-4.2-3b logo
granite-4.2-3b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-3B-Base Parameters 3B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Engli…

Open Source 3.0B ★ 9.0 ↓ 49.3K
LFM2.5-2.6B logo
LFM2.5-2.6B
LiquidAI

LFM2.5-2.6B is part of LFM2.5, a family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with a 128K context window and agentic post-training.

Open Source 2.6B ★ 8.0 ↓ 121.6K
LFM2.5-2.6B-Base logo
LFM2.5-2.6B-Base
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 2.6B ★ 8.0 ↓ 10.3K
ERNIE-4.5-300B-A47B-Base-PT logo
ERNIE-4.5-300B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 398
Step3-VL-10B-Base logo
Step3-VL-10B-Base
stepfun-ai

STEP3-VL-10B is a lightweight open-source foundation model designed to redefine the trade-off between compact efficiency and frontier-level multimodal intelligence. Despite its compact 10B parameter footprint , STEP3-VL-10B excels in visual perception , complex reasoning , and hu…

Multimodal 10.0B ★ 8.0 ↓ 335
ERNIE-4.5-300B-A47B-Base-Paddle logo
ERNIE-4.5-300B-A47B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 177
LFM2.5-8B-A1B logo
LFM2.5-8B-A1B
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment. It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 8.0B ★ 7.0 ↓ 25.8K