AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,115 models for "Chat" Compare
🤖
deepseek-coder-33b-instruct
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 33.0B ↓ 4.4K
🤖
GLM-4-32B-0414
zai-org

The GLM family welcomes new members, the GLM-4-32B-0414 series models, featuring 32 billion parameters. Its performance is comparable to OpenAI’s GPT series and DeepSeek’s V3/R1 series. It also supports very user-friendly local deployment features. GLM-4-32B-Base-0414 was pre-tra…

Open Source 32.0B ↓ 4.4K
🤖
command-a-plus-05-2026-w4a4
CohereLabs

Command A+ is an open source model with 25 billion active parameters and 218B total parameters model optimized for agentic, multilingual, and reasoning-heavy tasks with a focus on enterprise performance, while also providing support for vision inputs for processing image inputs.

Open Source ↓ 4.4K
🤖
Qwen3-Swallow-8B-RL-v0.2
tokyotech-llm

Qwen3-Swallow v0.2 is a family of large language models available in 8B , 30B-A3B , and 32B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifia…

Open Source 8.0B ↓ 4.3K
🤖
Hermes-3-Llama-3.2-3B
NousResearch

Hermes 3 3B is a small but mighty new addition to the Hermes series of LLMs by Nous Research, and is Nous's first fine-tune in this parameter class.

Open Source 3.0B ↓ 4.2K
🤖
MiniCPM-MoE-8x2B
openbmb

The MiniCPM-MoE-8x2B is a decoder-only transformer-based generative language model.

Open Source 2.0B ↓ 4.2K
🤖
falcon-40b-instruct
tiiuae

Falcon-40B-Instruct is a 40B parameters causal decoder-only model built by TII based on Falcon-40B and finetuned on a mixture of Baize. It is made available under the Apache 2.0 license.

Open Source 40.0B ↓ 4.2K
🤖
Seed-Coder-8B-Instruct
ByteDance-Seed

Introduction We are thrilled to introduce Seed-Coder, a powerful, transparent, and parameter-efficient family of open-source code models at the 8B scale, featuring base, instruct, and reasoning variants. Seed-Coder contributes to promote the evolution of open code models through…

Code 8.0B ↓ 4.1K
🤖
Llama-3.1-Swallow-8B-Instruct-v0.3
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 8.0B ↓ 4.1K
🤖
Falcon-H1-Tiny-R-0.6B
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 0.6B ↓ 4K
🤖
internlm2-chat-1_8b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 8.0B ↓ 4K
🤖
CapRL-Qwen3VL-4B
internlm

CapRL 📖 Paper 🏠 Github 🤗 CapRL Collection 🤗 Daily Paper

Multimodal 4.0B ↓ 4K