AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
InternVL3_5-4B logo
InternVL3_5-4B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.7B ↓ 19.4K
ERNIE-4.5-21B-A3B-PT logo
ERNIE-4.5-21B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 19.1K
EXAONE-3.0-7.8B-Instruct logo
EXAONE-3.0-7.8B-Instruct
LGAI-EXAONE

Open Source 7.8B ↓ 17.4K
InternVL3_5-4B-HF logo
InternVL3_5-4B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.7B ↓ 15.4K
ERNIE-4.5-0.3B-PT logo
ERNIE-4.5-0.3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 0.36B ↓ 14.9K
Aurora-Spec-Minimax-M2.5 logo
Aurora-Spec-Minimax-M2.5
togethercomputer

This is an EAGLE3 draft model trained from scratch (random initialization) using the Aurora inference-time training framework for speculative decoding. Unlike traditional approaches that fine-tune pre-trained models, this model is built entirely through Aurora's online training p…

Open Source ↓ 14.8K
Olmo-Hybrid-7B logo
Olmo-Hybrid-7B
allenai

We expand on our Olmo model series by introducing Olmo Hybrid, a new 7B hybrid RNN model in the Olmo family. Olmo Hybrid dramatically outperforms Olmo 3 in final performance, consistently showing roughly 2x data efficiency on core evals over the course of our pretraining run. We…

Open Source 7.0B ↓ 14.6K
LFM2.5-350M-Base logo
LFM2.5-350M-Base
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 0.35B ↓ 14.4K
InternVL3_5-1B-HF logo
InternVL3_5-1B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 1.1B ↓ 14.4K
LFM2.5-1.2B-Base logo
LFM2.5-1.2B-Base
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 1.2B ↓ 13.8K
Llama-3.1-Tulu-3-8B-SFT logo
Llama-3.1-Tulu-3-8B-SFT
allenai

Tülu3 is a leading instruction following model family, offering fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern post-training techniques. Tülu3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat…

Open Source 8.0B ↓ 13.5K
OLMo-2-1124-13B-Instruct logo
OLMo-2-1124-13B-Instruct
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 13.0B ↓ 12.5K