AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
SmolLM2-360M-Instruct logo
SmolLM2-360M-Instruct
HuggingFaceTB

1. Model Summary 2. Limitations 3. Training 4. License 5. Citation

Open Source ↓ 254.1K
SmolLM2-1.7B-Instruct logo
SmolLM2-1.7B-Instruct
HuggingFaceTB

1. Model Summary 2. Evaluation 3. Examples 4. Limitations 5. Training 6. License 7. Citation

Open Source 1.7B ↓ 220K
Phi-3-mini-128k-instruct logo
Phi-3-mini-128k-instruct
microsoft

🎉 Phi-4 : [multimodal-instruct onnx]; [mini-instruct onnx]

Open Source ↓ 160.7K
Phi-3.5-MoE-instruct logo
Phi-3.5-MoE-instruct
microsoft

Phi-3.5-MoE is a lightweight, state-of-the-art open model built upon datasets used for Phi-3 - synthetic data and filtered publicly available documents - with a focus on very high-quality, reasoning dense data. The model supports multilingual and comes with 128K context length (i…

Open Source ↓ 138.8K
OLMoE-1B-7B-0125-Instruct logo
OLMoE-1B-7B-0125-Instruct
allenai

OLMoE-1B-7B-0125-Instruct January 2025 is post-trained variant of the OLMoE-1B-7B January 2025 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset and further DPO training on this dataset, and finally RLVR training using this data. T…

Open Source 1.0B ↓ 134.8K
Kimi-Linear-48B-A3B-Base logo
Kimi-Linear-48B-A3B-Base
moonshotai

Kimi Linear: An Expressive, Efficient Attention Architecture

Open Source 48.0B ↓ 124.7K
OLMoE-1B-7B-0924 logo
OLMoE-1B-7B-0924
allenai

OLMoE-1B-7B is a Mixture-of-Experts LLM with 1B active and 7B total parameters released in September 2024 (0924). It yields state-of-the-art performance among models with a similar cost (1B) and is competitive with much larger models like Llama2-13B. OLMoE is 100% open-source.

Open Source 7.0B ↓ 110.4K
Llama-xLAM-2-8b-fc-r logo
Llama-xLAM-2-8b-fc-r
Salesforce

Large Action Models (LAMs) are advanced language models designed to enhance decision-making by translating user intentions into executable actions. As the brains of AI agents , LAMs autonomously plan and execute tasks to achieve specific goals, making them invaluable for automati…

Open Source 8.0B ↓ 95.3K
LFM2.5-230M logo
LFM2.5-230M
LiquidAI

LFM2.5 is a family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source ↓ 81.5K
InternVL2_5-4B-MPO logo
InternVL2_5-4B-MPO
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.0B ↓ 80.1K
Llama-3_3-Nemotron-Super-49B-v1_5 logo
Llama-3_3-Nemotron-Super-49B-v1_5
nvidia

Llama-3.3-Nemotron-Super-49B-v1.5 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reasoning,…

Reasoning 49.0B ↓ 73.6K
Hunyuan-A13B-Instruct logo
Hunyuan-A13B-Instruct
tencent

🤗  Hugging Face       🖥️  Official Website       🕖  HunyuanAPI       🕹️  Demo       🤖  ModelScope

Open Source 13.0B ↓ 72K