AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

351 models for "DPO" Compare
Kimi-K2-Thinking logo
Kimi-K2-Thinking
moonshotai

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, we built it as a thinking agent that reasons step-by-step while dynamically invoking tools. It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and oth…

Open Source ★ 17.0 ↓ 63.5K
granite-4.2-30b logo
granite-4.2-30b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-30B-Base Parameters 30B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Eng…

Open Source 30.0B ★ 15.0 ↓ 38.8K
Qwen3.5-9B logo
Qwen3.5-9B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 9.0B ★ 13.0 ↓ 8.5M
Qwen3.5-4B logo
Qwen3.5-4B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 4.0B ★ 13.0 ↓ 8.1M
NVIDIA-Nemotron-3-Super-120B-A12B-BF16 logo
NVIDIA-Nemotron-3-Super-120B-A12B-BF16
nvidia

:--- :--- Total Parameters 120B (12B active) Architecture LatentMoE - Mamba-2 + MoE + Attention hybrid with Multi-Token Prediction (MTP) Context Length Up to 1M tokens Minimum GPU Requirement 8× H100-80GB Supported Languages English, French, German, Italian, Japanese, Spanish, Ch…

Reasoning 120.0B ★ 13.0 ↓ 1M
NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4 logo
NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
nvidia

:--- :--- Total Parameters 120B (12B active) Architecture LatentMoE - Mamba-2 + MoE + Attention hybrid with Multi-Token Prediction (MTP) Context Length Up to 1M tokens Minimum GPU Requirement 1× B200 OR 1× DGX Spark Supported Languages English, French, German, Italian, Japanese,…

Reasoning 120.0B ★ 13.0 ↓ 678.6K
NVIDIA-Nemotron-3-Super-120B-A12B-FP8 logo
NVIDIA-Nemotron-3-Super-120B-A12B-FP8
nvidia

:--- :--- Total Parameters 120B (12B active) Architecture LatentMoE - Mamba-2 + MoE + Attention hybrid with Multi-Token Prediction (MTP) Context Length Up to 1M tokens Minimum GPU Requirement 2× H100-80GB Supported Languages English, French, German, Italian, Japanese, Spanish, Ch…

Reasoning 120.0B ★ 13.0 ↓ 64K
Llama-3_3-Nemotron-Super-49B-v1_5-FP8 logo
Llama-3_3-Nemotron-Super-49B-v1_5-FP8
nvidia

Llama-3.3-Nemotron-Super-49B-v1.5-FP8 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reason…

Reasoning 49.0B ★ 12.0 ↓ 76.7K
granite-4.2-8b logo
granite-4.2-8b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-8B-Base Parameters 8B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Engli…

Open Source 8.0B ★ 11.0 ↓ 133.6K
InternVL3_5-GPT-OSS-20B-A4B-Preview-HF logo
InternVL3_5-GPT-OSS-20B-A4B-Preview-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 21.2B ★ 10.0 ↓ 645.7K
InternVL3_5-GPT-OSS-20B-A4B-Preview logo
InternVL3_5-GPT-OSS-20B-A4B-Preview
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 20.0B ★ 10.0 ↓ 17.3K
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 1.5M