AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,295 models for "Transformer" Compare
🤖
Hermes-4-70B-FP8
NousResearch

Hermes 4 70B is a frontier, hybrid-mode reasoning model based on Llama-3.1-70B by Nous Research that is aligned to you .

Open Source 70.0B ★ 10.0 ↓ 3.3K
🤖
Falcon-H1R-7B
tiiuae

This repository presents Falcon-H1R-7B , a reasoning-specialized model introduced in the paper Falcon-H1R: Pushing the Reasoning Frontiers with a Hybrid Model for Efficient Test-Time Scaling.

Open Source 7.0B ★ 10.0 ↓ 2K
🤖
Falcon-H1R-7B-FP8
tiiuae

This repository presents post FP8 quantized Falcon-H1R-7B-FP8 via NVIDIA Model Optimizer, enabling efficient inference while preserving the strong reasoning introduced in the paper Falcon-H1R: Pushing the Reasoning Frontiers with a Hybrid Model for Efficient Test-Time Scaling.

Open Source 7.0B ★ 10.0 ↓ 228
🤖
Llama-3.3-70B-Instruct
meta-llama

Open Source 70.0B ★ 9.0 ↓ 549.3K
🤖
NVIDIA-Nemotron-Nano-9B-v2
nvidia

The pretraining data has a cutoff date of September 2024.

Open Source 9.0B ★ 9.0 ↓ 431.8K
🤖
NVIDIA-Nemotron-3-Nano-4B-BF16
nvidia

The pretraining data has a cutoff date of September 2024\.

Open Source 4.0B ★ 9.0 ↓ 377.9K
🤖
granite-4.1-30b
ibm-granite

Model Summary: Granite-4.1-30B is a 30B parameter long-context instruct model finetuned from Granite-4.1-30B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an i…

Open Source 30.0B ★ 9.0 ↓ 340.2K
🤖
NVIDIA-Nemotron-Nano-12B-v2-VL-FP8
nvidia

NVIDIA-Nemotron-Nano-VL-12B-V2-FP8 is the quantized version of the NVIDIA Nemotron Nano VL V2 model, which is an auto-regressive vision language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Nemotron Nano VL FP4 QAD mod…

Multimodal 12.0B ★ 9.0 ↓ 125.7K
🤖
NVIDIA-Nemotron-Nano-9B-v2-FP8
nvidia

The pretraining data has a cutoff date of September 2024.

Open Source 9.0B ★ 9.0 ↓ 121.3K
🤖
Step3-VL-10B
stepfun-ai

- 🚀 Online Demo : Explore Step3-VL-10B on Hugging Face Spaces ! - 📢 [Notice] FP8 Quantization Support : FP8 quantized weights are now available. (Download link) - 📢 [Notice] vLLM Support: vLLM integration is now officially supported! (PR 32329) - ✅ [Fixed] HF Inference: Resolved…

Multimodal 10.0B ★ 9.0 ↓ 36.4K
🤖
ERNIE-4.5-300B-A47B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 9.0 ↓ 2.4K
🤖
Hermes-4-405B
NousResearch

Hermes 4 405B is a frontier, hybrid-mode reasoning model based on Llama-3.1-405B by Nous Research that is aligned to you .

Open Source 405.0B ★ 9.0 ↓ 748