AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,115 models for "Chat" Compare
🤖
ERNIE-4.5-VL-28B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 28.0B ↓ 54.6K
🤖
LFM2.5-VL-450M
LiquidAI

LFM2.5‑VL-450M is Liquid AI's refreshed version of the first vision-language model, LFM2-VL-450M, built on an updated backbone LFM2.5-350M and tuned for stronger real-world performance. Find more about LFM2.5 family of models in our blog post.

Multimodal 0.45B ↓ 54.3K
🤖
LFM2.5-230M
LiquidAI

LFM2.5 is a family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source ↓ 52.1K
🤖
bloom-1b7
bigscience

BLOOM LM BigScience Large Open-science Open-access Multilingual Language Model Model Card

Open Source 1.7B ↓ 50.9K
🤖
Hunyuan-A13B-Instruct
tencent

🤗  Hugging Face       🖥️  Official Website       🕖  HunyuanAPI       🕹️  Demo       🤖  ModelScope

Open Source 13.0B ↓ 48.2K
🤖
MiniCPM4.1-8B
openbmb

GitHub Repo Technical Report Join Us 👋 Contact us in Discord and WeChat

Reasoning 8.0B ↓ 47.8K
🤖
InternVL3_5-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 8.5B ↓ 47.7K
🤖
MiniCPM-V-2_6
openbmb

Multimodal 8.0B ↓ 46.3K
🤖
Kimi-VL-A3B-Thinking
moonshotai

[!Warning] This model has a new version: Kimi-VL-A3B-Thinking-2506. Please consider using the new 2506 version for better abilties on general visual understanding, reasoning, video and agent scenarios.

Multimodal 16.0B ↓ 45.9K
🤖
granite-3.1-8b-instruct
ibm-granite

Model Summary: Granite-3.1-8B-Instruct is a 8B parameter long-context instruct model finetuned from Granite-3.1-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets tailored for solving long context pr…

Open Source 8.0B ↓ 45.3K
🤖
Llama-3.1-70B
meta-llama

Open Source 70.0B ↓ 45.2K
🤖
MiMo-7B-RL
XiaomiMiMo

━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ Unlocking the Reasoning Potential of Language Model From Pretraining to Posttraining ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━

Open Source 7.0B ↓ 44.8K