AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
LFM2.5-1.2B-JP-202606 logo
LFM2.5-1.2B-JP-202606
LiquidAI

LFM2.5-1.2B-JP-202606 is our latest general purpose Japanese chat model, delivering significant improvements in knowledge, instruction following, math, code, and tool-use over both the models of comparable size and LFM2.5-1.2B-JP. It sets a new benchmark for state-of-the-art perf…

Open Source 1.2B ↓ 9.9K
OLMo-2-0325-32B-Instruct logo
OLMo-2-0325-32B-Instruct
allenai

OLMo 2 32B Instruct March 2025 is post-trained variant of the OLMo-2 32B March 2025 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training on this dataset. Tülu 3 is designe…

Open Source 32.0B ↓ 9.9K
InternVL3_5-14B logo
InternVL3_5-14B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 15.1B ↓ 9.3K
Yi-34B-200K logo
Yi-34B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 9.2K
Yi-9B logo
Yi-9B
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 9.0B ↓ 9K
Nous-Hermes-2-Yi-34B logo
Nous-Hermes-2-Yi-34B
NousResearch

Nous Hermes 2 - Yi-34B is a state of the art Yi Fine-tune.

Open Source 34.0B ↓ 8.9K
Yi-9B-200K logo
Yi-9B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 9.0B ↓ 8.9K
Llama-3.1-Swallow-8B-Instruct-v0.5 logo
Llama-3.1-Swallow-8B-Instruct-v0.5
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 8.0B ↓ 8.6K
tulu-2-7b logo
tulu-2-7b
allenai

Tulu is a series of language models that are trained to act as helpful assistants. Tulu 2 7B is a fine-tuned version of Llama 2 that was trained on a mix of publicly available, synthetic and human datasets.

Open Source 7.0B ↓ 8.5K
OLMo-2-0425-1B-SFT logo
OLMo-2-0425-1B-SFT
allenai

OLMo 2 1B SFT April 2025 is post-trained variant of the allenai/OLMo-2-0425-1B model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset. Tülu 3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such…

Open Source 1.0B ↓ 8.4K
InternVL3_5-30B-A3B-Instruct logo
InternVL3_5-30B-A3B-Instruct
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 30.85B ↓ 7.3K
InternVL3_5-2B logo
InternVL3_5-2B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 2.3B ↓ 7.2K