AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,296 models for "Transformer" Compare
🤖
ERNIE-4.5-300B-A47B-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 349
🤖
Hermes-4-405B-FP8
NousResearch

Hermes 4 405B is a frontier, hybrid-mode reasoning model based on Llama-3.1-405B by Nous Research that is aligned to you .

Open Source 405.0B ★ 9.0 ↓ 305
🤖
Step3-VL-10B-FP8
stepfun-ai

- 🚀 Online Demo : Explore Step3-VL-10B on Hugging Face Spaces ! - 📢 [Notice] FP8 Quantization Support : FP8 quantized weights are now available. (Download link) - 📢 [Notice] vLLM Support: vLLM integration is now officially supported! (PR 32329) - ✅ [Fixed] HF Inference: Resolved…

Multimodal 10.0B ★ 9.0 ↓ 207
🤖
Step3-VL-10B-Base
stepfun-ai

STEP3-VL-10B is a lightweight open-source foundation model designed to redefine the trade-off between compact efficiency and frontier-level multimodal intelligence. Despite its compact 10B parameter footprint , STEP3-VL-10B excels in visual perception , complex reasoning , and hu…

Multimodal 10.0B ★ 9.0 ↓ 181
🤖
ERNIE-4.5-300B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 167
🤖
ERNIE-4.5-300B-A47B-W4A8C8-TP4-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 54
🤖
ERNIE-4.5-300B-A47B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 50
🤖
ERNIE-4.5-300B-A47B-FP8-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 37
🤖
Llama-3.1-405B-FP8
meta-llama

Open Source 405.0B ★ 8.0 ↓ 488.3K
🤖
Kimi-Linear-48B-A3B-Instruct
moonshotai

Kimi Linear: An Expressive, Efficient Attention Architecture

Open Source 48.0B ★ 8.0 ↓ 185K
🤖
Llama-3.1-405B
meta-llama

Open Source 405.0B ★ 8.0 ↓ 141.7K
🤖
LFM2.5-8B-A1B
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment. It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 8.0B ★ 8.0 ↓ 100K