AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,112 models for "Chat" Compare
🤖
InternVL2_5-4B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 3.7B ↓ 230.9K
🤖
DeepSeek-V3.1
deepseek-ai

DeepSeek-V3.1 is a hybrid model that supports both thinking mode and non-thinking mode. Compared to the previous version, this upgrade brings improvements in multiple aspects:

Open Source ↓ 228.6K
🤖
LLaDA2.0-mini
inclusionAI

LLaDA2.0-mini is a diffusion language model featuring a 16BA1B Mixture-of-Experts (MoE) architecture. As an enhanced, instruction-tuned iteration of the LLaDA series, it is optimized for practical applications.

Open Source ↓ 223.3K
🤖
GLM-4.1V-9B-Thinking
zai-org

📖 View the GLM-4.1V-9B-Thinking paper . 📍 Using GLM-4.1V-9B-Thinking API at Zhipu Foundation Model Open Platform

Open Source 9.0B ↓ 220.4K
🤖
Llama-3.1-Nemotron-Nano-8B-v1
nvidia

Llama-3.1-Nemotron-Nano-8B-v1 is a large language model (LLM) which is a derivative of Meta Llama-3.1-8B-Instruct (AKA the reference model). It is a reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling.

Reasoning 8.0B ↓ 218.1K
🤖
granite-guardian-3.3-8b
ibm-granite

Model Summary: Granite Guardian 3.3 8b is a specialized Granite 3.3 8B model designed to judge if the input prompts and the output responses of an LLM based system meet specified criteria. The model comes pre-baked with certain criteria including but not limited to: jailbreak att…

Open Source 8.0B ↓ 217.6K
🤖
olmOCR-2-7B-1025
allenai

Full BF16 version of olmOCR-2-7B-1025-FP8. We recommend using the FP8 version for all practical purposes except further fine tuning.

Multimodal 7.0B ↓ 213.8K
🤖
granite-docling-258M
ibm-granite

granite-docling-258m Granite Docling is a multimodal Image-Text-to-Text model engineered for efficient document conversion. It preserves the core features of Docling while maintaining seamless integration with DoclingDocuments to ensure full compatibility.

Multimodal 0.258B ↓ 212.6K
🤖
gemma-3n-E2B-it
google

Multimodal 2.0B ↓ 204.9K
🤖
DeepSeek-R1-0528
deepseek-ai

The DeepSeek R1 model has undergone a minor version upgrade, with the current version being DeepSeek-R1-0528. In the latest update, DeepSeek R1 has significantly improved its depth of reasoning and inference capabilities by leveraging increased computational resources and introdu…

Reasoning ↓ 204.2K
🤖
InternVL3-1B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 0.94B ↓ 200.8K
🤖
paligemma-3b-pt-224
google

Multimodal 3.0B ↓ 190.8K