AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
GLM-Z1-Rumination-32B-0414 logo
GLM-Z1-Rumination-32B-0414
zai-org

The GLM family welcomes a new generation of open-source models, the GLM-4-32B-0414 series, featuring 32 billion parameters. Its performance is comparable to OpenAI's GPT series and DeepSeek's V3/R1 series, and it supports very user-friendly local deployment features. GLM-4-32B-Ba…

Open Source 32.0B ↓ 670
Hunyuan-0.5B-Pretrain logo
Hunyuan-0.5B-Pretrain
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 0.5B ↓ 629
stablelm-base-alpha-7b logo
stablelm-base-alpha-7b
stabilityai

📢 DISCLAIMER : The StableLM-Base-Alpha models have been superseded. Find the latest versions in the Stable LM Collection here.

Open Source 7.0B ↓ 626
nanowhale-100m logo
nanowhale-100m
HuggingFaceTB

A small ~110M parameter language model implementing the DeepSeek-V4 architecture , fine-tuned for chat/instruction following. Trained from scratch — no weights from DeepSeek-V4 were used.

Open Source ↓ 613
ERNIE-4.5-21B-A3B-Base-PT logo
ERNIE-4.5-21B-A3B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 611
JustRL-II-base-model logo
JustRL-II-base-model
openbmb

This is the RL initialization checkpoint used in the blog JustRL II: Scaling Small LLMs to 128K Reasoning with a Critic (中文版).

Reasoning 2.5B ↓ 599
japanese-stablelm-instruct-ja_vocab-beta-7b logo
japanese-stablelm-instruct-ja_vocab-beta-7b
stabilityai

Japanese-StableLM-Instruct-JAVocab-Beta-7B

Open Source 7.0B ↓ 582
GLM-4.5-Base logo
GLM-4.5-Base
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.5 technical blog , technical report , and Zhipu AI technical documentation . 📍 Use GLM-4.5 API services on Z.ai API Platform (Global) or Zhipu AI Open Platform (Mainland China) . 👉 One click to GLM-4.5 .

Open Source ↓ 574
japanese-stablelm-base-ja_vocab-beta-7b logo
japanese-stablelm-base-ja_vocab-beta-7b
stabilityai

A cute robot wearing a kimono writes calligraphy with one single brush — Stable Diffusion XL

Open Source 7.0B ↓ 568
HiLS-Attention-7B logo
HiLS-Attention-7B
tencent

HiLS-Attention is a chunk-wise sparse attention mechanism that learns chunk selection end-to-end under the language-modeling loss, enabling native sparse training for efficient long-context modeling. This repository hosts the 7B checkpoint continued-trained on top of an OLMo3-sty…

Open Source 7.0B ↓ 568
MiMo-7B-RL-0530 logo
MiMo-7B-RL-0530
XiaomiMiMo

━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ Unlocking the Reasoning Potential of Language Model From Pretraining to Posttraining ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━

Open Source 7.0B ↓ 567
Hunyuan-4B-Pretrain logo
Hunyuan-4B-Pretrain
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 4.0B ↓ 560