AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,108 models for "Chat" Compare
🤖
granite-4.2-3b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-3B-Base Parameters 3B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Engli…

Open Source 3.0B ★ 14.0 ↓ 8.2K
🤖
Qwen: Qwen3 Next 80B A3B Instruct
qwen

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

Closed Source ★ 14.0
🤖
Qwen: Qwen3 Next 80B A3B Thinking
qwen

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

Closed Source ★ 14.0
🤖
diffusiongemma-26B-A4B-it
google

Hugging Face GitHub Launch Blog Documentation License : Apache 2.0 Authors : Google DeepMind

Open Source 26.0B ★ 13.0 ↓ 1.3M
🤖
diffusiongemma-26B-A4B-it-NVFP4
nvidia

--- pipeline tag: text-generation base model: google/diffusiongemma-26B-A4B-it license: apache-2.0 license name: apache-license-2.0 license link: https://ai.google.dev/gemma/apache 2 tags: - nvidia - ModelOpt - DiffusionGemma-26B-A4B-IT - quantized - NVFP4 - nvfp4 ---

Open Source 26.0B ★ 13.0 ↓ 299.1K
🤖
MiniCPM5-1B
openbmb

MiniCPM Tech Report MiniCPM Wiki(Chinese) GitHub Repo UltraData MiniCPM Desk Pet Online Demo

Open Source 1.0B ★ 12.0 ↓ 790.5K
🤖
Llama-3_3-Nemotron-Super-49B-v1_5-FP8
nvidia

Llama-3.3-Nemotron-Super-49B-v1.5-FP8 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reason…

Reasoning 49.0B ★ 12.0 ↓ 253.4K
🤖
MiniCPM5-1B-SFT
openbmb

MiniCPM Tech Report MiniCPM Wiki(Chinese) GitHub Repo UltraData MiniCPM Desk Pet Online Demo

Open Source 1.0B ★ 12.0 ↓ 15.3K
🤖
MiniCPM5-1B-Base
openbmb

MiniCPM Tech Report MiniCPM Wiki(Chinese) GitHub Repo UltraData MiniCPM Desk Pet Online Demo

Open Source 1.0B ★ 12.0 ↓ 11.4K
🤖
sarvam-1
sarvamai

Sarvam-1 is a 2-billion parameter language model specifically optimized for Indian languages. It provides best in-class performance in 10 Indic languages (bn, gu, hi, kn, ml, mr, or, pa, ta, te) when compared with popular models like Gemma-2-2B and Llama-3.2-3B. It is also compet…

Open Source ★ 12.0 ↓ 8.4K
🤖
sarvam-105b
sarvamai

Want a smaller model? Download Sarvam-30B!

Open Source 105.0B ★ 12.0 ↓ 6.5K
🤖
LFM2.5-2.6B
LiquidAI

LFM2.5-2.6B is part of LFM2.5, a family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with a 128K context window and agentic post-training.

Open Source 2.6B ★ 11.0 ↓ 198.3K