AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 logo
NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4
nvidia

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4

Open Source 75.0B ↓ 306.2K
DeepSeek-R1-Distill-Qwen-7B logo
DeepSeek-R1-Distill-Qwen-7B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 7.0B ↓ 292.8K
granite-docling-258M logo
granite-docling-258M
ibm-granite

granite-docling-258m Granite Docling is a multimodal Image-Text-to-Text model engineered for efficient document conversion. It preserves the core features of Docling while maintaining seamless integration with DoclingDocuments to ensure full compatibility.

Multimodal 0.258B ↓ 276K
Kimi-K2.5-NVFP4 logo
Kimi-K2.5-NVFP4
nvidia

Description: The NVIDIA Kimi-K2.5-NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2.5 model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2.5 NVFP4 model is qu…

Open Source ↓ 268.4K
paligemma-3b-pt-224 logo
paligemma-3b-pt-224
google

Multimodal 3.0B ↓ 235.5K
LLaDA2.0-mini logo
LLaDA2.0-mini
inclusionAI

LLaDA2.0-mini is a diffusion language model featuring a 16BA1B Mixture-of-Experts (MoE) architecture. As an enhanced, instruction-tuned iteration of the LLaDA series, it is optimized for practical applications.

Open Source ↓ 215.5K
LocateAnything-3B logo
LocateAnything-3B
nvidia

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

Multimodal 3.0B ↓ 210.3K
Llama-Guard-3-8B logo
Llama-Guard-3-8B
meta-llama

Open Source 8.0B ↓ 194.7K
Florence-2-base-ft logo
Florence-2-base-ft
microsoft

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Multimodal 0.23B ↓ 193.3K
DeepSeek-R1-Distill-Llama-8B logo
DeepSeek-R1-Distill-Llama-8B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 8.0B ↓ 187.7K
granite-vision-4.1-4b logo
granite-vision-4.1-4b
ibm-granite

Model Summary: Granite Vision 4.1 4B is a vision-language model (VLM) that delivers frontier-level performance on structured document extraction tasks — chart extraction, table extraction, and semantic key-value pair extraction — in a compact 4B parameter footprint, providing a l…

Multimodal 4.0B ↓ 179.4K
Phi-3-mini-128k-instruct logo
Phi-3-mini-128k-instruct
microsoft

🎉 Phi-4 : [multimodal-instruct onnx]; [mini-instruct onnx]

Open Source ↓ 160.7K