AI Agent Hub

LLM Models · Qwen

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

104 models from Qwen Compare
Qwen2-1.5B-Instruct logo
Qwen2-1.5B-Instruct
Qwen

Qwen2 is the new series of Qwen large language models. For Qwen2, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters, including a Mixture-of-Experts model. This repo contains the instruction-tuned 1.5B Qwen2…

open-source 1.5B ↓ 949.6K
Qwen2-0.5B logo
Qwen2-0.5B
Qwen

Qwen2 is the new series of Qwen large language models. For Qwen2, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters, including a Mixture-of-Experts model. This repo contains the 0.5B Qwen2 base language mod…

open-source 0.5B ↓ 880.5K
Qwen2.5-Math-1.5B logo
Qwen2.5-Math-1.5B
Qwen

[!Warning] 🚨 Qwen2.5-Math mainly supports solving English and Chinese math problems through CoT and TIR. We do not recommend using this series of models for other tasks.

reasoning 1.5B ↓ 859K
Qwen3-VL-235B-A22B-Instruct logo
Qwen3-VL-235B-A22B-Instruct
Qwen

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

multimodal 235.0B ↓ 825.5K
Qwen3-Coder-30B-A3B-Instruct logo
Qwen3-Coder-30B-A3B-Instruct
Qwen

Qwen3-Coder is available in multiple sizes. Today, we're excited to introduce Qwen3-Coder-30B-A3B-Instruct . This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements:

code 30.0B ↓ 774.8K
Qwen2.5-Coder-1.5B-Instruct logo
Qwen2.5-Coder-1.5B-Instruct
Qwen

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). As of now, Qwen2.5-Coder has covered six mainstream model sizes, 0.5, 1.5, 3, 7, 14, 32 billion parameters, to meet the needs of different developers. Qwen2.5-Coder brings…

code 1.5B ↓ 773K
Qwen: Qwen3 Coder 30B A3B Instruct logo
Qwen: Qwen3 Coder 30B A3B Instruct
qwen

Qwen3-Coder-30B-A3B-Instruct is a 30.5B parameter Mixture-of-Experts (MoE) model with 128 experts (8 active per forward pass), designed for advanced code generation, repository-scale understanding, and agentic tool use. Built on the...

code
Qwen2.5 72B Instruct logo
Qwen2.5 72B Instruct
qwen

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

closed-source
Qwen: Qwen-Plus logo
Qwen: Qwen-Plus
qwen

Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.

closed-source
Qwen: Qwen2.5 VL 72B Instruct logo
Qwen: Qwen2.5 VL 72B Instruct
qwen

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

multimodal
Qwen: Qwen3 235B A22B logo
Qwen: Qwen3 235B A22B
qwen

Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

closed-source
Qwen: Qwen3 235B A22B Instruct 2507 logo
Qwen: Qwen3 235B A22B Instruct 2507
qwen

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following,...

closed-source