AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

562 models for "Vision" Compare
Qwen3.6-35B-A3B logo
Qwen3.6-35B-A3B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 35.0B ★ 18.0 ↓ 3.5M
Qwen: Qwen3.5-122B-A10B logo
Qwen: Qwen3.5-122B-A10B
qwen

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...

Multimodal 122.0B ★ 18.0
gemma-4-26B-A4B-it logo
gemma-4-26B-A4B-it
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 26.0B ★ 17.0 ↓ 12.3M
Gemma-4-26B-A4B-NVFP4 logo
Gemma-4-26B-A4B-NVFP4
nvidia

Description: Gemma 4 26B IT is an open multimodal model built by Google DeepMind that handles text and image inputs, can process video as sequences of frames, and generates text output. It is designed to deliver frontier-level performance for reasoning, agentic workflows, coding,…

Open Source 26.0B ★ 17.0 ↓ 922K
gemma-4-26B-A4B-it-qat-q4_0-unquantized logo
gemma-4-26B-A4B-it-qat-q4_0-unquantized
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 26.0B ★ 17.0 ↓ 125.1K
gemma-4-31B-it logo
gemma-4-31B-it
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 15.0 ↓ 9.6M
gemma-4-31B logo
gemma-4-31B
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 15.0 ↓ 555K
gemma-4-31B-it-qat-q4_0-unquantized logo
gemma-4-31B-it-qat-q4_0-unquantized
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 15.0 ↓ 315.7K
gemma-4-31B-it-qat-w4a16-ct logo
gemma-4-31B-it-qat-w4a16-ct
google

Hugging Face GitHub Launch Blog Documentation Technical Report License : Apache 2.0 Authors : Google DeepMind

Open Source 31.0B ★ 15.0 ↓ 249.5K
granite-4.2-30b logo
granite-4.2-30b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-30B-Base Parameters 30B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Eng…

Open Source 30.0B ★ 15.0 ↓ 38.8K
Llama-4-Maverick-17B-128E-Instruct-FP8 logo
Llama-4-Maverick-17B-128E-Instruct-FP8
meta-llama

Multimodal 400.0B ★ 14.0 ↓ 70.3K
Qwen3.5-9B logo
Qwen3.5-9B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 9.0B ★ 13.0 ↓ 8.5M