AI Agent Hub

LLM Models · Qwen

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

104 models from Qwen Compare
Qwen3.6-35B-A3B-FP8 logo
Qwen3.6-35B-A3B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

open-source 35.0B ★ 32.0 ↓ 13.3M
Qwen3.6-35B-A3B logo
Qwen3.6-35B-A3B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

open-source 35.0B ★ 32.0 ↓ 4.9M
Qwen3.5-35B-A3B logo
Qwen3.5-35B-A3B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

open-source 35.0B ★ 24.0 ↓ 2.5M
Qwen3.5-35B-A3B-FP8 logo
Qwen3.5-35B-A3B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

multimodal 35.0B ★ 24.0 ↓ 1.4M
Qwen3.5-9B logo
Qwen3.5-9B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

open-source 9.0B ★ 22.0 ↓ 12.6M
Qwen: Qwen3.5-9B (batch) logo
Qwen: Qwen3.5-9B (batch)
qwen

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

closed-source ★ 22.0
Qwen3-Coder-Next-FP8 logo
Qwen3-Coder-Next-FP8
Qwen

Today, we're announcing Qwen3-Coder-Next-FP8 , an open-weight language model designed specifically for coding agents and local development. It features the following key enhancements:

code 80.0B ★ 21.0 ↓ 1.8M
Qwen: Qwen3 Coder 480B A35B logo
Qwen: Qwen3 Coder 480B A35B
qwen

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

code ★ 21.0
Qwen: Qwen3 Coder Next logo
Qwen: Qwen3 Coder Next
qwen

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

code ★ 21.0
Qwen3.5-4B logo
Qwen3.5-4B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

open-source 4.0B ★ 20.0 ↓ 7.4M
Qwen: Qwen3 Next 80B A3B Instruct logo
Qwen: Qwen3 Next 80B A3B Instruct
qwen

Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...

closed-source ★ 14.0
Qwen: Qwen3 Next 80B A3B Thinking logo
Qwen: Qwen3 Next 80B A3B Thinking
qwen

Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...

closed-source ★ 14.0