AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,299 models for "Transformer" Compare
🤖
Qwen3-VL-4B-Instruct
Qwen

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Multimodal 4.0B ↓ 4M
🤖
Qwen2.5-7B-Instruct-AWQ
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 7.0B ↓ 3.9M
🤖
Qwen2.5-VL-3B-Instruct
Qwen

--- license name: qwen-research license link: https://huggingface.co/Qwen/Qwen2.5-VL-3B-Instruct/blob/main/LICENSE language: - en pipeline tag: image-text-to-text tags: - multimodal library name: transformers ---

Multimodal 3.0B ↓ 3.8M
🤖
Qwen3-1.7B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 1.7B ↓ 3.7M
🤖
Qwen3-4B-Instruct-2507
Qwen

We introduce the updated version of the Qwen3-4B non-thinking mode , named Qwen3-4B-Instruct-2507 , featuring the following key enhancements:

Open Source 4.0B ↓ 3.5M
🤖
Qwen-72B
Qwen

🤗 Hugging Face &nbsp&nbsp &nbsp&nbsp🤖 ModelScope &nbsp&nbsp &nbsp&nbsp 📑 Paper &nbsp&nbsp | &nbsp&nbsp🖥️ Demo WeChat (微信) &nbsp&nbsp &nbsp&nbsp Discord &nbsp&nbsp | &nbsp&nbsp API

Open Source 72.0B ↓ 3.2M
🤖
gemma-3-1b-it
google

Open Source 1.0B ↓ 3.2M
🤖
Unlimited-OCR
baidu

Welcome the Era of One-shot Long-horizon Parsing.

Multimodal 3.336B ↓ 3.1M
🤖
Qwen3-VL-2B-Instruct
Qwen

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Multimodal 2.0B ↓ 2.9M
🤖
Florence-2-base
microsoft

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Multimodal 0.23B ↓ 2.8M
🤖
Qwen2.5-14B-Instruct
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 14.7B ↓ 2.7M
🤖
SmolLM2-135M
HuggingFaceTB

1. Model Summary 2. Limitations 3. Training 4. License 5. Citation

Open Source ↓ 2.5M