AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
aya-expanse-32b logo
aya-expanse-32b
CohereLabs

Open Source 32.0B ↓ 7K
InternVL2_5-1B-MPO logo
InternVL2_5-1B-MPO
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 0.9B ↓ 7K
xLAM-1b-fc-r logo
xLAM-1b-fc-r
Salesforce

[Homepage] [APIGen Paper] [ActionStudio Paper] [Discord] [Dataset] [Github]

Open Source 1.35B ↓ 6.5K
InternVL3_5-38B logo
InternVL3_5-38B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.4B ↓ 6.5K
Llama-3.1-Tulu-3-8B logo
Llama-3.1-Tulu-3-8B
allenai

Tülu 3 is a leading instruction following model family, offering a post-training package with fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern techniques. This is one step of a bigger process to training fully open-source models, lik…

Open Source 8.0B ↓ 6.3K
OLMo-2-1124-7B-SFT logo
OLMo-2-1124-7B-SFT
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 6.2K
OLMo-2-0325-32B logo
OLMo-2-0325-32B
allenai

We introduce OLMo 2 32B, the largest model in the OLMo 2 family. OLMo 2 was pre-trained on OLMo-mix-1124 and uses Dolmino-mix-1124 for mid-training.

Open Source 32.0B ↓ 5.7K
LFM2-700M logo
LFM2-700M
LiquidAI

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.

Open Source 0.7B ↓ 5.7K
Llama-3.1-Swallow-8B-Instruct-v0.3 logo
Llama-3.1-Swallow-8B-Instruct-v0.3
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 8.0B ↓ 5.3K
InternVL3_5-2B-HF logo
InternVL3_5-2B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 2.3B ↓ 5.2K
Hermes-3-Llama-3.2-3B logo
Hermes-3-Llama-3.2-3B
NousResearch

Hermes 3 3B is a small but mighty new addition to the Hermes series of LLMs by Nous Research, and is Nous's first fine-tune in this parameter class.

Open Source 3.0B ↓ 4.6K
Hermes-2-Pro-Llama-3-8B logo
Hermes-2-Pro-Llama-3-8B
NousResearch

Hermes 2 Pro is an upgraded, retrained version of Nous Hermes 2, consisting of an updated and cleaned version of the OpenHermes 2.5 Dataset, as well as a newly introduced Function Calling and JSON Mode dataset developed in-house.

Open Source 8.0B ↓ 4K