AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,112 models for "Chat" Compare
🤖
Kimi-K2.5
moonshotai

📰   Tech Blog         📄   Paper

Open Source ↓ 554.9K
🤖
Cosmos-Reason2-8B
nvidia

Multimodal 8.0B ↓ 527K
🤖
DeepSeek-R1-Distill-Qwen-1.5B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 1.5B ↓ 488.7K
🤖
MiniMax-M2.5
MiniMaxAI

Join Our 💬 WeChat 🧩 Discord community. MiniMax Agent ⚡️ API MCP MiniMax Website 🤗 Hugging Face 🚀 Hugging Face API 🐙 GitHub 🤖️ ModelScope 📄 License: Modified-MIT

Open Source ↓ 487.4K
🤖
Llama-3.1-70B-Instruct
meta-llama

Open Source 70.0B ↓ 468.6K
🤖
bloom-560m
bigscience

BLOOM LM BigScience Large Open-science Open-access Multilingual Language Model Model Card

Open Source 0.56B ↓ 467.9K
🤖
blip2-opt-2.7b
Salesforce

BLIP-2 model, leveraging OPT-2.7b (a large language model with 2.7 billion parameters). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Open Source 2.7B ↓ 462.5K
🤖
DeepSeek-R1-Distill-Qwen-14B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 14.0B ↓ 460.1K
🤖
gemma-3-27b-it
google

Multimodal 27.0B ↓ 441.3K
🤖
Qwen3-8B-FP8
nvidia

Description: The NVIDIA Qwen3-8B FP8 model is the quantized version of Alibaba's Qwen3-8B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3-8B FP8 model is quantized with Te…

Open Source 8.0B ↓ 437.5K
🤖
NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4
nvidia

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4

Open Source 75.0B ↓ 437.2K
🤖
GLM-5.1-FP8
zai-org

👋 Join our WeChat or Discord community. 📖 Check out the GLM-5.1 blog and GLM-5 Technical report . 📍 Use GLM-5.1 API services on Z.ai API Platform. 🔜 GLM-5.1 will be available on chat.z.ai in the coming days.

Open Source ↓ 398K