AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

562 models for "Vision" Compare
MiMo-V2-Flash-Base logo
MiMo-V2-Flash-Base
XiaomiMiMo

🤗 HuggingFace   📔 Technical Report   📰 Blog   Play around!   🗨️ Xiaomi MiMo Studio   🎨 Xiaomi MiMo API Platform

Open Source ★ 25.0 ↓ 873
Qwen3.5-35B-A3B logo
Qwen3.5-35B-A3B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 35.0B ★ 24.0 ↓ 1.5M
Qwen3.5-35B-A3B-FP8 logo
Qwen3.5-35B-A3B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

Multimodal 35.0B ★ 24.0 ↓ 1.3M
Qwen: Qwen3.5-9B (batch) logo
Qwen: Qwen3.5-9B (batch)
qwen

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...

Closed Source ★ 22.0
Qwen: Qwen3.5 397B A17B logo
Qwen: Qwen3.5 397B A17B
qwen

The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...

Closed Source ★ 21.0
Qwen: Qwen3.5 Plus 2026-02-15 logo
Qwen: Qwen3.5 Plus 2026-02-15
qwen

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...

Closed Source ★ 21.0
Olmo-3-1025-7B logo
Olmo-3-1025-7B
allenai

We introduce Olmo 3, a new family of 7B and 32B models. This suite includes Base, Instruct, and Think variants. The Base models were trained using a staged training approach.

Open Source 7.0B ★ 20.0 ↓ 127.5K
Olmo-3-32B-Think-SFT logo
Olmo-3-32B-Think-SFT
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 20.0 ↓ 43.5K
Olmo-3-1125-32B logo
Olmo-3-1125-32B
allenai

We introduce Olmo 3, a new family of 7B and 32B models. This suite includes Base, Instruct, and Think variants. The Base models were trained using a staged training approach.

Open Source 32.0B ★ 20.0 ↓ 21.9K
Olmo-3-32B-Think logo
Olmo-3-32B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 20.0 ↓ 14.6K
Qwen3.6-35B-A3B-NVFP4 logo
Qwen3.6-35B-A3B-NVFP4
nvidia

Description: The NVIDIA Qwen3.6-35B-A3B-NVFP4 model is the quantized version of Alibaba's Qwen3.6-35B-A3B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.6-35B-A3B-NVFP4 m…

Open Source 35.0B ★ 18.0 ↓ 5.3M
Qwen3.6-35B-A3B-FP8 logo
Qwen3.6-35B-A3B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

Open Source 35.0B ★ 18.0 ↓ 4.5M