AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,296 models for "Transformer" Compare
🤖
Olmo-3.1-32B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 8.0 ↓ 5.7K
🤖
Hermes-3-Llama-3.1-405B
NousResearch

Hermes 3 405B is the latest flagship model in the Hermes series of LLMs by Nous Research, and the first full parameter finetune since the release of Llama-3.1 405B.

Open Source 405.0B ★ 8.0 ↓ 599
🤖
Ring-flash-2.0
inclusionAI

This model is presented in the paper Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model.

Open Source ★ 8.0 ↓ 225
🤖
Qwen3.5-2B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. In light of its parameter scale, the intende…

Open Source 2.0B ★ 7.0 ↓ 2.8M
🤖
Qwen3.5-2B-Base
Qwen

[!Note] This repository contains model weights and configuration files for the pre-trained only model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, etc. The intended use cases are fine-tuning, in-context lear…

Open Source 2.0B ★ 7.0 ↓ 1.4M
🤖
Ministral-3-3B-Instruct-2512-ONNX
mistralai

[!Tip] This model was contributed by Xenova from Hugging Face. We sincerely appreciate the integration and community collaboration. While preliminary functionality checks have been performed, comprehensive testing has not yet been completed. We recommend you to proceed with cauti…

Open Source 3.0B ★ 7.0 ↓ 404
🤖
granite-4.1-8b
ibm-granite

Model Summary: Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from Granite-4.1-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an impr…

Open Source 8.0B ★ 6.0 ↓ 1.1M
🤖
Phi-4-mini-instruct
microsoft

🎉 Phi-4 : [mini-reasoning reasoning] [multimodal-instruct onnx]; [mini-instruct onnx]

Open Source ★ 6.0 ↓ 453.2K
🤖
Llama-3.2-90B-Vision-Instruct
meta-llama

Multimodal 90.0B ★ 6.0 ↓ 103.8K
🤖
Olmo-3.1-32B-Instruct
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 6.0 ↓ 16.4K
🤖
sarvam-30b
sarvamai

Want a bigger model? Download Sarvam-105B!

Open Source 30.0B ★ 6.0 ↓ 14.7K
🤖
granite-4.1-8b-base
ibm-granite

Model Summary: Granite‑4.1‑8B‑Base is a decoder‑only language model with long‑context capabilities, designed to support a broad range of text‑to‑text generation tasks. In addition to standard generation, it supports Fill‑in‑the‑Middle (FIM) code completion through specialized pre…

Open Source 8.0B ★ 6.0 ↓ 6.1K