AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,115 models for "Chat" Compare
🤖
MiniCPM4-0.5B
openbmb

GitHub Repo Technical Report 👋 Join us on Discord and WeChat

Open Source 0.5B ↓ 43.6K
🤖
Molmo2-O-7B
allenai

Molmo2 is a family of open vision-language models developed by the Allen Institute for AI (Ai2) that support image, video and multi-image understanding and grounding. Molmo2 models are trained on publicly available third party datasets as referenced in our technical report and Mo…

Multimodal 7.0B ↓ 43.5K
🤖
Emu3-Chat-hf
BAAI

Emu3: Next-Token Prediction is All You Need

Multimodal 8.0B ↓ 41.7K
🤖
InternVL2-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 8.1B ↓ 41.6K
🤖
Meta-Llama-3.1-8B
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 8.0B ↓ 41.1K
🤖
granite-3.3-2b-instruct
ibm-granite

Model Summary: Granite-3.3-2B-Instruct is a 2-billion parameter 128K context length language model fine-tuned for improved reasoning and instruction-following capabilities. Built on top of Granite-3.3-2B-Base, the model delivers significant gains on benchmarks for measuring gener…

Open Source 2.0B ↓ 40K
🤖
command-a-plus-05-2026-bf16
CohereLabs

Command A+ is an open source model with 25 billion active parameters and 218B total parameters model optimized for agentic, multilingual, and reasoning-heavy tasks with a focus on enterprise performance, while also providing support for vision inputs for processing image inputs.

Open Source 218.0B ↓ 39.9K
🤖
academic-ds-9B
ByteDance-Seed

This is a 9B model whose architecture is deepseek-v3, trained from scratch using 350B+ tokens from fully open-source, English-only datasets. It is designed for development and debugging purposes within the open-source community.

Open Source 9.0B ↓ 39.3K
🤖
Kimi-K2-Instruct-0905
moonshotai

📰   Tech Blog         📄   Paper

Open Source ↓ 38.4K
🤖
MiniCPM-V-2_6-int4
openbmb

[2025.01.14] 🔥🔥 We open source MiniCPM-o 2.6 , with significant performance improvement over MiniCPM-V 2.6 , and support real-time speech-to-speech conversation and multimodal live streaming. Try it now.

Multimodal 8.0B ↓ 38.2K
🤖
MiniCPM4-8B
openbmb

GitHub Repo Technical Report Join Us 👋 Contact us in Discord and WeChat

Open Source 8.0B ↓ 37.9K
🤖
step3
stepfun-ai

📰   Step3 Model Blog         📄   Step3 System Blog

Open Source ↓ 37.7K