AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
ERNIE-4.5-VL-424B-A47B-Base-PT logo
ERNIE-4.5-VL-424B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 840
Llama-3.1-Swallow-8B-Instruct-v0.1 logo
Llama-3.1-Swallow-8B-Instruct-v0.1
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 8.0B ↓ 820
AI21-Jamba2-Mini logo
AI21-Jamba2-Mini
ai21labs

Jamba2 Mini is an open source small language model built for enterprise reliability. With 12B active parameters (52B total), it delivers precise question answering without the computational overhead of reasoning models. The model's SSM-Transformer architecture provides a memory-e…

Open Source ↓ 809
UI-Mate-9B logo
UI-Mate-9B
tencent

UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations

Open Source 9.0B ↓ 808
Ring-1T logo
Ring-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 Experience Now

Open Source ↓ 794
Hunyuan-4B-Instruct logo
Hunyuan-4B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 4.0B ↓ 770
UI-Mate-27B logo
UI-Mate-27B
tencent

UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations

Open Source 27.0B ↓ 761
LLaDA-UI logo
LLaDA-UI
inclusionAI

Bringing Block-wise Diffusion to Vision-Language GUI Agents

Multimodal 16.7B ↓ 760
ERNIE-4.5-VL-28B-A3B-Thinking logo
ERNIE-4.5-VL-28B-A3B-Thinking
baidu

🚀 Introducing ERNIE-4.5-VL-28B-A3B-Thinking: A Breakthrough in Multimodal AI

Multimodal 28.0B ↓ 735
InternVL3_5-4B-Instruct logo
InternVL3_5-4B-Instruct
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.7B ↓ 716
ERNIE-4.5-VL-424B-A47B-PT logo
ERNIE-4.5-VL-424B-A47B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 685
stablelm-2-12b-chat logo
stablelm-2-12b-chat
stabilityai

Stable LM 2 12B Chat is a 12 billion parameter instruction tuned language model trained on a mix of publicly available datasets and synthetic datasets, utilizing Direct Preference Optimization (DPO).

Open Source 12.0B ↓ 676