AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,297 models for "Transformer" Compare
🤖
MiniCPM-V-4.6-Thinking-BNB
openbmb

This repository hosts the bitsandbytes (NF4, 4-bit) quantized version of MiniCPM-V 4.6 Thinking. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6-Thinking.

Multimodal 1.3B ★ 4.0 ↓ 38.5K
🤖
MiniCPM-V-4.6-AWQ
openbmb

This repository hosts the AWQ (W4A16, AutoAWQ) quantized version of MiniCPM-V 4.6. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6.

Multimodal 1.3B ★ 4.0 ↓ 36.6K
🤖
MiniCPM-V-4.6-GPTQ
openbmb

This repository hosts the GPTQ (W4A16, GPTQModel) quantized version of MiniCPM-V 4.6. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6.

Multimodal 1.3B ★ 4.0 ↓ 31.3K
🤖
MiniCPM-V-4.6-Thinking-AWQ
openbmb

This repository hosts the AWQ (W4A16, AutoAWQ) quantized version of MiniCPM-V 4.6 Thinking. For the original BF16 weights and the full model card, please refer to openbmb/MiniCPM-V-4.6-Thinking.

Multimodal 1.3B ★ 4.0 ↓ 28.4K
🤖
Olmo-3-7B-Think-SFT
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 4.0 ↓ 14.1K
🤖
AI21-Jamba-Reasoning-3B
ai21labs

AI21’s Jamba Reasoning 3B is a top-performing reasoning model that packs leading scores on intelligence benchmarks and highly-efficient processing into a compact 3B build. Read the full blog post here.

Open Source 3.0B ★ 4.0 ↓ 13.5K
🤖
granite-4.1-3b-base
ibm-granite

Model Summary: Granite‑4.1‑3B‑Base is a decoder‑only language model with long‑context capabilities, designed to support a broad range of general text‑to‑text generation tasks, as well as fill‑in‑the‑Middle (FIM) code completion. This model shares the same underlying architecture…

Open Source 3.0B ★ 4.0 ↓ 7.2K
🤖
Olmo-3-7B-Think-DPO
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 4.0 ↓ 5K
🤖
Llama-3.2-11B-Vision-Instruct
meta-llama

Multimodal 11.0B ★ 3.0 ↓ 107.2K
🤖
Molmo-7B-D-0924
allenai

Molmo is a family of open vision-language models developed by the Allen Institute for AI. Molmo models are trained on PixMo, a dataset of 1 million, highly-curated image-text pairs. It has state-of-the-art performance among multimodal models with a similar size while being fully…

Multimodal 8.0B ★ 3.0 ↓ 23.6K
🤖
EXAONE-4.0-1.2B
LGAI-EXAONE

🎉 License Updated! We are pleased to announce our more flexible licensing terms 🤗 ✈️ Try on FriendliAI (licensed under commercial purposes) 📢 EXAONE 4.0 is officially supported by HuggingFace transformers! Please check out the guide below

Reasoning 1.2B ★ 3.0 ↓ 20.4K
🤖
OLMo-7B
allenai

For transformers versions v4.40.0 or newer, we suggest using OLMo 7B HF instead.

Open Source 7.0B ★ 3.0 ↓ 4.9K