AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
EXAONE-4.0-1.2B-GPTQ-Int8 logo
EXAONE-4.0-1.2B-GPTQ-Int8
LGAI-EXAONE

🎉 License Updated! We are pleased to announce our more flexible licensing terms 🤗 ✈️ Try on FriendliAI (licensed under commercial purposes) 📢 EXAONE 4.0 is officially supported by HuggingFace transformers! Please check out the guide below

Open Source 1.2B ★ 5.0 ↓ 128
Mistral: Ministral 3 8B 2512 (batch) logo
Mistral: Ministral 3 8B 2512 (batch)
mistralai

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

Multimodal 8.0B ★ 5.0
AI21-Jamba2-3B logo
AI21-Jamba2-3B
ai21labs

Introduction Jamba2 3B is an ultra-compact open source model designed to bring enterprise-grade reliability to on-device deployments. At just 3B parameters, it runs efficiently on consumer devices—iPhones, Androids, Macs, and PCs—while maintaining the grounding and instruction-fo…

Open Source 3.0B ★ 4.0 ↓ 26.9K
LFM2-2.6B logo
LFM2-2.6B
LiquidAI

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

Open Source 2.6B ★ 2.0 ↓ 11.4K
LFM2-8B-A1B logo
LFM2-8B-A1B
LiquidAI

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

Open Source 8.3B ★ 1.0 ↓ 17.7K
Qwen3-0.6B logo
Qwen3-0.6B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 0.6B ↓ 31.2M
Qwen3-8B logo
Qwen3-8B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 8.0B ↓ 10.1M
Qwen3-4B logo
Qwen3-4B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Reasoning 4.0B ↓ 9.6M
Llama-3.2-1B-Instruct logo
Llama-3.2-1B-Instruct
meta-llama

Open Source 1.23B ↓ 6.2M
Qwen3-14B logo
Qwen3-14B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 14.8B ↓ 4.8M
Qwen3-32B logo
Qwen3-32B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 32.0B ↓ 3.7M
Qwen3-4B-Instruct-2507 logo
Qwen3-4B-Instruct-2507
Qwen

We introduce the updated version of the Qwen3-4B non-thinking mode , named Qwen3-4B-Instruct-2507 , featuring the following key enhancements:

Open Source 4.0B ↓ 3.7M