AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
Yi-34B-Chat logo
Yi-34B-Chat
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 34.5K
Yi-6B-Chat logo
Yi-6B-Chat
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 32.5K
InternVL3_5-1B logo
InternVL3_5-1B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 1.1B ↓ 30K
ERNIE-4.5-VL-28B-A3B-PT logo
ERNIE-4.5-VL-28B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 28.0B ↓ 29.2K
LFM2-350M logo
LFM2-350M
LiquidAI

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

Open Source 0.35B ↓ 28.8K
SmolLM-135M-Instruct logo
SmolLM-135M-Instruct
HuggingFaceTB

SmolLM is a series of small language models available in three sizes: 135M, 360M, and 1.7B parameters.

Open Source 0.135B ↓ 26.9K
InternVL3_5-8B-HF logo
InternVL3_5-8B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 8.5B ↓ 26.5K
MiniCPM-2B-sft-bf16 logo
MiniCPM-2B-sft-bf16
openbmb

MiniCPM 技术报告 Technical Report OmniLMM 多模态模型 Multi-modal Model CPM-C 千亿模型试用 ~100B Model Trial

Open Source 2.4B ↓ 23.7K
Yi-6B-200K logo
Yi-6B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 22.9K
aya-expanse-8b logo
aya-expanse-8b
CohereLabs

Open Source 8.0B ↓ 20.4K
MiMo-V2.6-Distill-Qwen-9B logo
MiMo-V2.6-Distill-Qwen-9B
XiaomiMiMo

MiMo-V2.6-Distill-Qwen-9B is a 9B agentic model developed by Xiaomi MiMo through supervised fine-tuning of Qwen3.5-9B on MiMo-generated data. It covers coding, general-purpose agent tasks, visual coding, and cybersecurity. We release this SFT checkpoint as a starting point for op…

Open Source 9.0B ↓ 19.8K
OLMoE-1B-7B-0924-Instruct logo
OLMoE-1B-7B-0924-Instruct
allenai

OLMoE-1B-7B-Instruct is a Mixture-of-Experts LLM with 1B active and 7B total parameters released in September 2024 (0924) that has been adapted via SFT and DPO from OLMoE-1B-7B. It yields state-of-the-art performance among models with a similar cost (1B) and is competitive with m…

Open Source 1.0B ↓ 19.4K