AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
LFM2-1.2B logo
LFM2-1.2B
LiquidAI

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

Open Source 1.2B ↓ 68.2K
Kimi-K2-Instruct-0905 logo
Kimi-K2-Instruct-0905
moonshotai

📰   Tech Blog         📄   Paper

Open Source ↓ 64.7K
LFM2.5-350M logo
LFM2.5-350M
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source ↓ 63.5K
OLMo-2-0425-1B-Instruct logo
OLMo-2-0425-1B-Instruct
allenai

OLMo 2 1B Instruct April 2025 is post-trained variant of the allenai/OLMo-2-0425-1B-RLVR1 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training on this dataset. Tülu 3 is d…

Open Source 1.0B ↓ 60.8K
Phi-4-mini-reasoning logo
Phi-4-mini-reasoning
microsoft

Phi-4-mini-reasoning is a lightweight open model built upon synthetic data with a focus on high-quality, reasoning dense data further finetuned for more advanced math reasoning capabilities. The model belongs to the Phi-4 model family and supports 128K token context length.

Open Source ↓ 56.2K
Hunyuan-7B-Instruct logo
Hunyuan-7B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 7.0B ↓ 51.2K
Llama-3.2-1B logo
Llama-3.2-1B
NousResearch

The Meta Llama 3.2 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction-tuned generative models in 1B and 3B sizes (text in/text out). The Llama 3.2 instruction-tuned text only models are optimized for multilingual dialogue use cas…

Open Source 1.0B ↓ 48.9K
Hermes-2-Pro-Mistral-7B logo
Hermes-2-Pro-Mistral-7B
NousResearch

Hermes 2 Pro on Mistral 7B is the new flagship 7B Hermes!

Open Source 7.0B ↓ 48K
SmolLM2-1.7B logo
SmolLM2-1.7B
HuggingFaceTB

1. Model Summary 2. Evaluation 3. Limitations 4. Training 5. License 6. Citation

Open Source 1.7B ↓ 45.7K
InternVL3_5-8B logo
InternVL3_5-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 8.5B ↓ 40.6K
OLMo-2-1124-7B-Instruct logo
OLMo-2-1124-7B-Instruct
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 36.1K
SOLAR-10.7B-Instruct-v1.0 logo
SOLAR-10.7B-Instruct-v1.0
upstage

Meet 10.7B Solar: Elevating Performance with Upstage Depth UP Scaling!

Reasoning 10.7B ↓ 35.7K