AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,116 models for "Chat" Compare
🤖
Yi-1.5-6B
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 6.0B ↓ 12.9K
🤖
falcon-40b
tiiuae

Falcon-40B is a 40B parameters causal decoder-only model built by TII and trained on 1,000B tokens of RefinedWeb enhanced with curated corpora. It is made available under the Apache 2.0 license.

Open Source 40.0B ↓ 12.9K
🤖
falcon-mamba-7b-instruct
tiiuae

Model card for FalconMamba Instruct model

Open Source 7.0B ↓ 12.8K
🤖
InternVL3-14B-hf
OpenGVLab

InternVL3-14B Transformers 🤗 Implementation

Multimodal 15.1B ↓ 12.4K
🤖
codegeex4-all-9b
zai-org

CodeGeeX4: Open Multilingual Code Generation Model

Code 9.0B ↓ 12.3K
🤖
InternVL2-4B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 4.2B ↓ 12.2K
🤖
EXAONE-3.0-7.8B-Instruct
LGAI-EXAONE

Open Source 7.8B ↓ 12K
🤖
Jamba-v0.1
ai21labs

This is the base version of the Jamba model. We’ve since released a better, instruct-tuned version, Jamba-1.5-Mini. For even greater performance, check out the scaled-up Jamba-1.5-Large.

Open Source ↓ 11.9K
🤖
glm-4-9b
zai-org

2024/08/12, 本仓库代码已更新并使用 transformers =4.44.0 , 请及时更新依赖。

Open Source 9.0B ↓ 11.9K
🤖
InternVL3-2B-hf
OpenGVLab

InternVL3-2B Transformers 🤗 Implementation

Multimodal 2.09B ↓ 11.7K
🤖
LFM2.5-350M-Base
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 0.35B ↓ 11.7K
🤖
OLMo-2-0425-1B-DPO
allenai

OLMo 2 1B DPO April 2025 is post-trained variant of the allenai/OLMo-2-0425-1B-SFT model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset and further DPO training on this dataset. Tülu 3 is designed for state-of-the-art performance on a…

Open Source 1.0B ↓ 11.5K