AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
RedPajama-INCITE-Chat-3B-v1
togethercomputer

RedPajama-INCITE-Chat-3B-v1 was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research re…

Open Source 3.0B ↓ 601
🤖
Falcon-H1-1.5B-Deep-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 1.5B ↓ 600
🤖
DeepHermes-3-Llama-3-3B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 3.0B ↓ 596
🤖
open-calm-7b
cyberagent

OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.

Open Source 7.0B ↓ 583
🤖
SEA-LION-v1-7B
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The size of the models range from 3 billion to 7 billion parameters. This is the card for the SEA-LION 7B base model.

Open Source 7.0B ↓ 578
🤖
Hunyuan-1.8B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 1.8B ↓ 575
🤖
calm2-7b
cyberagent

CyberAgentLM2 is a decoder-only language model pre-trained on the 1.3T tokens of publicly available Japanese and English datasets.

Open Source 7.0B ↓ 563
🤖
Qwen3-Swallow-30B-A3B-RL-v0.2-AWQ-INT4
tokyotech-llm

Qwen3-Swallow v0.2 is a family of large language models available in 8B , 30B-A3B , and 32B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifia…

Open Source 30.0B ↓ 562
🤖
VISTA-9B
inclusionAI

VISTA-9B are GUI-grounding vision-language models trained from Qwen3.5 9B backbones with VISTA: View-Consistent Self-Verified Training for GUI Grounding .

Open Source 9.0B ↓ 558
🤖
codegen2-1B_P
Salesforce

CodeGen2 is a family of autoregressive language models for program synthesis , introduced in the paper:

Code 1.0B ↓ 557
🤖
LLaDA2.1-flash
inclusionAI

🚀 LLaDA2.1-flash is now live on ZenmuxAI ! Try it via API 🛠️ or Chat 💬: https://zenmux.ai/inclusionai/llada2.1-flash

Open Source ↓ 548
🤖
Ling-2.6-flash-base
inclusionAI

🤗 Hugging Face      🤖 ModelScope       Tech Report       💻 GitHub

Open Source ↓ 545