AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,109 models for "Chat" Compare
🤖
DeepHermes-3-Mistral-24B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 24.0B ★ 5.0 ↓ 636
🤖
MiniCPM-V-4.6
openbmb

A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone

Multimodal 1.3B ★ 4.0 ↓ 560.5K
🤖
granite-4.1-3b
ibm-granite

Model Summary: Granite-4.1-3B is a 3B parameter long-context instruct model finetuned from Granite-4.1-3B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an impr…

Open Source 3.0B ★ 4.0 ↓ 177.7K
🤖
MiniCPM-V-4.6-Thinking
openbmb

A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone

Multimodal 1.3B ★ 4.0 ↓ 90.1K
🤖
MiniCPM-V-4
openbmb

A GPT-4V Level MLLM for Single Image, Multi Image and Video on Your Phone

Multimodal 4.1B ★ 4.0 ↓ 87.3K
🤖
Olmo-3-7B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 4.0 ↓ 71.5K
🤖
MiniCPM-V-4.6-BNB
openbmb

This repository hosts the bitsandbytes (NF4, 4-bit) quantized version of MiniCPM-V 4.6. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6.

Multimodal 1.3B ★ 4.0 ↓ 45.4K
🤖
MiniCPM-V-4.6-Thinking-GPTQ
openbmb

This repository hosts the GPTQ (W4A16, GPTQModel) quantized version of MiniCPM-V 4.6 Thinking. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6-Thinking.

Multimodal 1.3B ★ 4.0 ↓ 39.2K
🤖
MiniCPM-V-4.6-Thinking-BNB
openbmb

This repository hosts the bitsandbytes (NF4, 4-bit) quantized version of MiniCPM-V 4.6 Thinking. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6-Thinking.

Multimodal 1.3B ★ 4.0 ↓ 38.5K
🤖
MiniCPM-V-4.6-AWQ
openbmb

This repository hosts the AWQ (W4A16, AutoAWQ) quantized version of MiniCPM-V 4.6. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6.

Multimodal 1.3B ★ 4.0 ↓ 36.6K
🤖
MiniCPM-V-4.6-GPTQ
openbmb

This repository hosts the GPTQ (W4A16, GPTQModel) quantized version of MiniCPM-V 4.6. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6.

Multimodal 1.3B ★ 4.0 ↓ 31.3K
🤖
MiniCPM-V-4.6-Thinking-AWQ
openbmb

This repository hosts the AWQ (W4A16, AutoAWQ) quantized version of MiniCPM-V 4.6 Thinking. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6-Thinking.

Multimodal 1.3B ★ 4.0 ↓ 28.4K