AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
MiniCPM-V-4_5-GPTQ
openbmb

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

Open Source ↓ 372
🤖
Llama-SEA-LION-v3-8B
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 8.0B ↓ 371
🤖
HiLS-Attention-7B
tencent

HiLS-Attention is a chunk-wise sparse attention mechanism that learns chunk selection end-to-end under the language-modeling loss, enabling native sparse training for efficient long-context modeling. This repository hosts the 7B checkpoint continued-trained on top of an OLMo3-sty…

Open Source 7.0B ↓ 365
🤖
cogagent-chat-hf
zai-org

🔥 News : The new version CogAgent-9B-20241220 has been released! Welcome to visit CogAgent GitHub and Technical Report to explore and use our latest model.

Open Source ↓ 364
🤖
Llama-SEA-LION-v2-8B-IT
aisingapore

Llama-SEA-LION-v2-8B-IT SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 8.0B ↓ 363
🤖
ZwZ-4B
inclusionAI

ZwZ-4B is a fine-grained multimodal perception model built upon Qwen3-VL-4B. It is trained using Region-to-Image Distillation (R2I) combined with reinforcement learning, enabling superior fine-grained visual understanding in a single forward pass — no inference-time zooming or to…

Open Source 4.0B ↓ 362
🤖
Ling-2.5-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 Experience Link Coming Soon~

Open Source ↓ 357
🤖
stablelm-base-alpha-7b-v2
stabilityai

StableLM-Base-Alpha-7B-v2 is a 7 billion parameter decoder-only language model pre-trained on diverse English datasets. This model is the successor to the first StableLM-Base-Alpha-7B model, addressing previous shortcomings through the use of improved data sources and mixture rat…

Open Source 7.0B ↓ 357
🤖
SmolLM3-3B-GSM8K-SFT
HuggingFaceTB

Fine-tuned version of HuggingFaceTB/SmolLM3-3B-Base optimized for grade school math (GSM8K benchmark).

Open Source 3.0B ↓ 355
🤖
bloomz-7b1-mt
bigscience

1. Model Summary 2. Use 3. Limitations 4. Training 5. Evaluation 7. Citation

Open Source ↓ 353
🤖
bloom-1b1-intermediate
bigscience

WARNING: The checkpoints on this repo are not fully trained model. Evaluations of intermediary checkpoints and the final model will be added when conducted (see below).

Open Source ↓ 353
🤖
k2-merged-3.5T-bf16
NousResearch

Experimental frankenmerge of Kimi K2-07, 09 and Base

Open Source ↓ 351