AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
Qwen3-1.7B logo
Qwen3-1.7B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 1.7B ↓ 3.1M
Qwen3-14B-AWQ logo
Qwen3-14B-AWQ
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Reasoning 14.8B ↓ 3M
Qwen3-8B-AWQ logo
Qwen3-8B-AWQ
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 8.0B ↓ 2.3M
Qwen3.5-27B logo
Qwen3.5-27B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 27.0B ↓ 1.9M
SmolLM2-135M logo
SmolLM2-135M
HuggingFaceTB

1. Model Summary 2. Limitations 3. Training 4. License 5. Citation

Open Source ↓ 1.7M
SmolLM2-135M-Instruct logo
SmolLM2-135M-Instruct
HuggingFaceTB

1. Model Summary 2. Limitations 3. Training 4. License 5. Citation

Open Source ↓ 1.7M
Qwen3-30B-A3B logo
Qwen3-30B-A3B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Reasoning 30.5B ↓ 1.4M
Qwen3-4B-Instruct-2507-FP8 logo
Qwen3-4B-Instruct-2507-FP8
Qwen

We introduce the updated version of the Qwen3-4B-FP8 non-thinking mode , named Qwen3-4B-Instruct-2507-FP8 , featuring the following key enhancements:

Open Source 4.0B ↓ 1.1M
Qwen3-32B-AWQ logo
Qwen3-32B-AWQ
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 32.8B ↓ 1M
Qwen3-30B-A3B-Instruct-2507 logo
Qwen3-30B-A3B-Instruct-2507
Qwen

We introduce the updated version of the Qwen3-30B-A3B non-thinking mode , named Qwen3-30B-A3B-Instruct-2507 , featuring the following key enhancements:

Open Source 30.0B ↓ 953.2K
MiniMax-M2.7 logo
MiniMax-M2.7
MiniMaxAI

Join Our 💬 WeChat 🧩 Discord community. MiniMax Agent ⚡️ API CLI MiniMax Website 🤗 Hugging Face 🐙 GitHub 🤖️ ModelScope 📄 LICENSE

Open Source ↓ 942.2K
HunyuanOCR logo
HunyuanOCR
tencent

HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better

Multimodal 1.0B ↓ 939.5K