AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,315 models for "Transformer" Compare
🤖
GLM-4-32B-0414
zai-org

The GLM family welcomes new members, the GLM-4-32B-0414 series models, featuring 32 billion parameters. Its performance is comparable to OpenAI’s GPT series and DeepSeek’s V3/R1 series. It also supports very user-friendly local deployment features. GLM-4-32B-Base-0414 was pre-tra…

Open Source 32.0B ↓ 4.4K
🤖
command-a-plus-05-2026-w4a4
CohereLabs

Command A+ is an open source model with 25 billion active parameters and 218B total parameters model optimized for agentic, multilingual, and reasoning-heavy tasks with a focus on enterprise performance, while also providing support for vision inputs for processing image inputs.

Open Source ↓ 4.4K
🤖
Molmo-72B-0924
allenai

Molmo is a family of open vision-language models developed by the Allen Institute for AI. Molmo models are trained on PixMo, a dataset of 1 million, highly-curated image-text pairs. It has state-of-the-art performance among multimodal models with a similar size while being fully…

Open Source 72.0B ↓ 4.2K
🤖
falcon-rw-1b
tiiuae

Falcon-RW-1B is a 1B parameters causal decoder-only model built by TII and trained on 350B tokens of RefinedWeb. It is made available under the Apache 2.0 license.

Open Source 1.0B ↓ 4.2K
🤖
Hermes-3-Llama-3.2-3B
NousResearch

Hermes 3 3B is a small but mighty new addition to the Hermes series of LLMs by Nous Research, and is Nous's first fine-tune in this parameter class.

Open Source 3.0B ↓ 4.2K
🤖
MiniCPM-MoE-8x2B
openbmb

The MiniCPM-MoE-8x2B is a decoder-only transformer-based generative language model.

Open Source 2.0B ↓ 4.2K
🤖
falcon-40b-instruct
tiiuae

Falcon-40B-Instruct is a 40B parameters causal decoder-only model built by TII based on Falcon-40B and finetuned on a mixture of Baize. It is made available under the Apache 2.0 license.

Open Source 40.0B ↓ 4.2K
🤖
Seed-Coder-8B-Instruct
ByteDance-Seed

Introduction We are thrilled to introduce Seed-Coder, a powerful, transparent, and parameter-efficient family of open-source code models at the 8B scale, featuring base, instruct, and reasoning variants. Seed-Coder contributes to promote the evolution of open code models through…

Code 8.0B ↓ 4.1K
🤖
Llama-3.1-Swallow-8B-Instruct-v0.3
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 8.0B ↓ 4.1K
🤖
Falcon-H1-Tiny-R-0.6B
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 0.6B ↓ 4K
🤖
internlm2-chat-1_8b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 8.0B ↓ 4K
🤖
tiny-random-stablelm-2
stabilityai

This repository stores a development version of Stable LM 2 for sanity-checking/debugging the transformers implementation.

Open Source ↓ 4K