AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Kimi-K2-Thinking-NVFP4 logo
Kimi-K2-Thinking-NVFP4
nvidia

Description: The NVIDIA Kimi-K2-Thinking-NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2-Thinking model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2-Think…

Reasoning 1000.0B ★ 11.0 ↓ 166K
Qwen3-Next-80B-A3B-Instruct-NVFP4 logo
Qwen3-Next-80B-A3B-Instruct-NVFP4
nvidia

Description: The NVIDIA Qwen3-Next-80B-A3B-Instruct NVFP4 model is the quantized version of Alibaba's Qwen3-Next-80B-A3B-Instruct model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA…

Open Source 80.0B ★ 11.0 ↓ 85.5K
gpt-oss-20b logo
gpt-oss-20b
openai

Try gpt-oss · Guides · Model card · OpenAI blog

Open Source 20.0B ★ 10.0 ↓ 6.1M
diffusiongemma-26B-A4B-it-NVFP4 logo
diffusiongemma-26B-A4B-it-NVFP4
nvidia

--- pipeline tag: text-generation base model: google/diffusiongemma-26B-A4B-it license: apache-2.0 license name: apache-license-2.0 license link: https://ai.google.dev/gemma/apache 2 tags: - nvidia - ModelOpt - DiffusionGemma-26B-A4B-IT - quantized - NVFP4 - nvfp4 ---

Open Source 26.0B ★ 10.0 ↓ 92.2K
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 1.5M
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 905.2K
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 246K
NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16
nvidia

NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16

Open Source 30.0B ★ 9.0 ↓ 79.2K
ERNIE-4.5-300B-A47B-PT logo
ERNIE-4.5-300B-A47B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 2.8K
ERNIE-4.5-300B-A47B-Paddle logo
ERNIE-4.5-300B-A47B-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 472
ERNIE-4.5-300B-A47B-Base-PT logo
ERNIE-4.5-300B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 398
ERNIE-4.5-300B-A47B-2Bits-Paddle logo
ERNIE-4.5-300B-A47B-2Bits-Paddle
baidu

The advanced capabilities of the ERNIE 4.5 models, particularly the MoE-based A47B and A3B series, are underpinned by several key technical innovations:

Open Source 300.0B ★ 8.0 ↓ 199