AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
Ling-3.0-flash-base-30T logo
Ling-3.0-flash-base-30T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 OpenRouter   

Open Source 124.0B ★ 20.0 ↓ 387
Ling-3.0-flash-base-midtrain logo
Ling-3.0-flash-base-midtrain
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 OpenRouter   

Open Source 124.0B ★ 20.0 ↓ 323
qwen2.5-bakeneko-32b-instruct logo
qwen2.5-bakeneko-32b-instruct
rinna

Qwen2.5 Bakeneko 32B Instruct (rinna/qwen2.5-bakeneko-32b-instruct)

Open Source 32.0B ★ 20.0 ↓ 144
deepseek-r1-distill-qwen2.5-bakeneko-32b logo
deepseek-r1-distill-qwen2.5-bakeneko-32b
rinna

DeepSeek R1 Distill Qwen2.5 Bakeneko 32B (rinna/deepseek-r1-distill-qwen2.5-bakeneko-32b)

Reasoning 32.0B ★ 20.0 ↓ 137
qwq-bakeneko-32b logo
qwq-bakeneko-32b
rinna

QwQ Bakeneko 32B (rinna/qwq-bakeneko-32b)

Open Source 32.0B ★ 20.0 ↓ 128
inclusionAI: Ling 3.0 Flash Sante logo
inclusionAI: Ling 3.0 Flash Sante
inclusionai

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

Closed Source 124.0B ★ 20.0
Qwen3.6-35B-A3B-NVFP4 logo
Qwen3.6-35B-A3B-NVFP4
nvidia

Description: The NVIDIA Qwen3.6-35B-A3B-NVFP4 model is the quantized version of Alibaba's Qwen3.6-35B-A3B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.6-35B-A3B-NVFP4 m…

Open Source 35.0B ★ 18.0 ↓ 5.3M
Qwen3.5-122B-A10B-NVFP4 logo
Qwen3.5-122B-A10B-NVFP4
nvidia

Description: The NVIDIA Qwen3.5-122B-A10B-NVFP4 model is the quantized version of Alibaba's Qwen3.5-122B-A10B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.5-122B-A10B N…

Open Source 122.0B ★ 18.0 ↓ 1M
Gemma-4-26B-A4B-NVFP4 logo
Gemma-4-26B-A4B-NVFP4
nvidia

Description: Gemma 4 26B IT is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output. It is designed to deliver frontier-level performance for reasoning, agentic workflows, coding,…

Open Source 26.0B ★ 17.0 ↓ 922K
Gemma-4-31B-IT-NVFP4 logo
Gemma-4-31B-IT-NVFP4
nvidia

Description: Gemma 4 31B IT is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output. It is designed to deliver frontier-level performance for reasoning, agentic workflows, coding,…

Open Source 31.0B ★ 15.0 ↓ 1.4M
granite-4.2-30b logo
granite-4.2-30b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-30B-Base Parameters 30B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Eng…

Open Source 30.0B ★ 15.0 ↓ 38.8K
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 logo
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4
nvidia

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4

Open Source 30.0B ★ 13.0 ↓ 839.9K