AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
Falcon-H1-7B-Instruct-GPTQ-Int4
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 7.0B ↓ 213
🤖
AHN-Mamba2-for-Qwen-2.5-Instruct-3B
ByteDance-Seed

AHN: Artificial Hippocampus Networks for Efficient Long-Context Modeling

Open Source 3.0B ↓ 212
🤖
Spatial-SSRL-Qwen3VL-4B
internlm

📖 Paper 🏠 Github 🤗 Spatial-SSRL-7B Model 🤗 Spatial-SSRL-3B Model 🤗 Spatial-SSRL-Qwen3VL-4B Model 🤗 Spatial-SSRL-81k Dataset 📰 Daily Paper

Multimodal 4.0B ↓ 211
🤖
codegen25-7b-instruct_P
Salesforce

Authors: Erik Nijkamp\ , Hiroaki Hayashi\ , Yingbo Zhou, Caiming Xiong

Code 7.0B ↓ 211
🤖
Baichuan2-13B-Chat-4bits
baichuan-inc

🦉GitHub 💬WeChat 百川API支持搜索增强和192K长窗口,新增百川搜索增强知识库、限时免费! 🚀 百川大模型在线对话平台 已正式向公众开放 🎉

Open Source 13.0B ↓ 210
🤖
starcoder-co-target
bigcode

An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 80

Code ↓ 210
🤖
Qianfan-VL-3B
baidu

Qianfan-VL: Domain-Enhanced Universal Vision-Language Models

Multimodal 3.0B ↓ 209
🤖
c4ai-command-r-v01-4bit
CohereLabs

Open Source ↓ 207
🤖
starcoder-co-format
bigcode

An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 50

Code ↓ 207
🤖
starcoder-o
bigcode

An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 45

Code ↓ 207
🤖
bloom-7b1-petals
bigscience

This model is a version of bigscience/bloom-7b1 post-processed to be run at home using the Petals swarm.

Open Source ↓ 206
🤖
GTA1-7B
Salesforce

Reinforcement learning (RL) (e.g., GRPO) helps with grounding because of its inherent objective alignment—rewarding successful clicks—rather than encouraging long textual Chain-of-Thought (CoT) reasoning. Unlike approaches that rely heavily on verbose CoT reasoning, GRPO directly…

Open Source 7.0B ↓ 204