AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

562 models for "Vision" Compare
InternVL3_5-GPT-OSS-20B-A4B-Preview-HF logo
InternVL3_5-GPT-OSS-20B-A4B-Preview-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 21.2B ★ 10.0 ↓ 645.7K
diffusiongemma-26B-A4B-it logo
diffusiongemma-26B-A4B-it
google

Hugging Face GitHub Launch Blog Documentation License : Apache 2.0 Authors : Google DeepMind

Open Source 26.0B ★ 10.0 ↓ 627.1K
Llama-4-Scout-17B-16E-Instruct logo
Llama-4-Scout-17B-16E-Instruct
meta-llama

Multimodal 109.0B ★ 10.0 ↓ 174.9K
InternVL3_5-GPT-OSS-20B-A4B-Preview logo
InternVL3_5-GPT-OSS-20B-A4B-Preview
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 20.0B ★ 10.0 ↓ 17.3K
Qwen3-Coder-Next-FP8 logo
Qwen3-Coder-Next-FP8
Qwen

Today, we're announcing Qwen3-Coder-Next-FP8 , an open-weight language model designed specifically for coding agents and local development. It features the following key enhancements:

Code 80.0B ★ 9.0 ↓ 1.1M
Qwen3-Coder-Next logo
Qwen3-Coder-Next
Qwen

Today, we're announcing Qwen3-Coder-Next , an open-weight language model designed specifically for coding agents and local development. It features the following key enhancements:

Code 80.0B ★ 9.0 ↓ 599.2K
granite-4.1-30b logo
granite-4.1-30b
ibm-granite

Model Summary: Granite-4.1-30B is a 30B parameter long-context instruct model finetuned from Granite-4.1-30B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an i…

Open Source 30.0B ★ 9.0 ↓ 254.1K
NVIDIA-Nemotron-Nano-12B-v2-VL-FP8 logo
NVIDIA-Nemotron-Nano-12B-v2-VL-FP8
nvidia

NVIDIA-Nemotron-Nano-VL-12B-V2-FP8 is the quantized version of the NVIDIA Nemotron Nano VL V2 model, which is an auto-regressive vision language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Nemotron Nano VL FP4 QAD mod…

Multimodal 12.0B ★ 9.0 ↓ 125.2K
granite-4.2-3b logo
granite-4.2-3b
ibm-granite

--- --- Developers Granite Team, IBM Model Type Decoder-only Dense Transformer (Reasoning) Architecture GraniteForCausalLM Base Model Granite-4.1-3B-Base Parameters 3B Context Length Natively Supports 128K (Long-context extension to 512K) Precision bfloat16 Tested Languages Engli…

Open Source 3.0B ★ 9.0 ↓ 49.3K
Mistral: Mistral Large 3 2512 (batch) logo
Mistral: Mistral Large 3 2512 (batch)
mistralai

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Open Source 675.0B ★ 9.0
Step3-VL-10B logo
Step3-VL-10B
stepfun-ai

- 🚀 Online Demo : Explore Step3-VL-10B on Hugging Face Spaces ! - 📢 [Notice] FP8 Quantization Support : FP8 quantized weights are now available. (Download link) - 📢 [Notice] vLLM Support: vLLM integration is now officially supported! (PR 32329) - ✅ [Fixed] HF Inference: Resolved…

Multimodal 10.0B ★ 8.0 ↓ 26.9K
ERNIE-4.5-300B-A47B-PT logo
ERNIE-4.5-300B-A47B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 2.8K