AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,110 models for "Chat" Compare
🤖
Qwen2.5-Math-1.5B
Qwen

[!Warning] 🚨 Qwen2.5-Math mainly supports solving English and Chinese math problems through CoT and TIR. We do not recommend using this series of models for other tasks.

Reasoning 1.5B ↓ 859K
🤖
Qwen3-VL-235B-A22B-Instruct
Qwen

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Multimodal 235.0B ↓ 825.5K
🤖
GLM-5-FP8
zai-org

👋 Join our WeChat or Discord community. 📖 Check out the GLM-5 Technical report . 📍 Use GLM-5 API services on Z.ai API Platform. 👉 One click to GLM-5 .

Open Source ↓ 825.1K
🤖
deepseek-coder-7b-instruct-v1.5
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 7.0B ↓ 812.8K
🤖
Llama-2-7b-hf
meta-llama

Open Source 7.0B ↓ 812.4K
🤖
MiniMax-M2.5-NVFP4
nvidia

Description: The NVIDIA MiniMax-M2.5-NVFP4 model is the quantized version of MiniMax's MiniMax-M2.5 model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA MiniMax-M2.5 NVFP4 model is q…

Open Source ↓ 776.5K
🤖
Qwen3-Coder-30B-A3B-Instruct
Qwen

Qwen3-Coder is available in multiple sizes. Today, we're excited to introduce Qwen3-Coder-30B-A3B-Instruct . This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements:

Code 30.0B ↓ 774.8K
🤖
Qwen2.5-Coder-1.5B-Instruct
Qwen

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). As of now, Qwen2.5-Coder has covered six mainstream model sizes, 0.5, 1.5, 3, 7, 14, 32 billion parameters, to meet the needs of different developers. Qwen2.5-Coder brings…

Code 1.5B ↓ 773K
🤖
gemma-2-2b-it
google

Open Source 2.0B ↓ 764.6K
🤖
DeepSeek-Coder-V2-Lite-Instruct
deepseek-ai

DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence

Code ↓ 717.8K
🤖
GOT-OCR2_0
stepfun-ai

General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

Multimodal 0.58B ↓ 701.6K
🤖
gemma-2-9b-it
google

Open Source 9.0B ↓ 693.9K