AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,109 models for "Chat" Compare
🤖
gemma-3-1b-it
google

Open Source 1.0B ↓ 3.2M
🤖
Unlimited-OCR
baidu

Welcome the Era of One-shot Long-horizon Parsing.

Multimodal 3.336B ↓ 3.1M
🤖
Qwen3-VL-2B-Instruct
Qwen

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Multimodal 2.0B ↓ 2.9M
🤖
Qwen2.5-14B-Instruct
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 14.7B ↓ 2.7M
🤖
Qwen3.5-27B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 27.0B ↓ 2.5M
🤖
Qwen3-30B-A3B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Reasoning 30.5B ↓ 2.4M
🤖
Qwen2.5-Coder-7B-Instruct
Qwen

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). As of now, Qwen2.5-Coder has covered six mainstream model sizes, 0.5, 1.5, 3, 7, 14, 32 billion parameters, to meet the needs of different developers. Qwen2.5-Coder brings…

Code 7.6B ↓ 2.4M
🤖
Qwen3-14B-AWQ
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Reasoning 14.8B ↓ 2.3M
🤖
Qwen3-VL-8B-Instruct-FP8
Qwen

This repository contains an FP8 quantized version of the Qwen3-VL-8B-Instruct model. The quantization method is fine-grained fp8 quantization with block size of 128, and its performance metrics are nearly identical to those of the original BF16 model. Enjoy!

Multimodal 8.0B ↓ 2.3M
🤖
GLM-OCR
zai-org

👋 Join our WeChat and Discord community 📍 Use GLM-OCR's API 👉 GLM-OCR SDK Recommended 📖 Technical Report

Multimodal 0.9B ↓ 2.1M
🤖
Qwen2.5-14B-Instruct-AWQ
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 14.7B ↓ 2.1M
🤖
Qwen2.5-32B-Instruct
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 32.0B ↓ 2M