AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

562 models for "Vision" Compare
DeepSeek: DeepSeek Flash Latest logo
DeepSeek: DeepSeek Flash Latest
~deepseek

This model always redirects to the latest model in the DeepSeek Flash family.

Multimodal 552.0B ★ 39.0
Qwen3.6-27B-FP8 logo
Qwen3.6-27B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

Open Source 27.0B ★ 38.0 ↓ 2.4M
Qwen3.6-27B logo
Qwen3.6-27B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 27.0B ★ 38.0 ↓ 2.3M
Qwen3.6-27B-NVFP4 logo
Qwen3.6-27B-NVFP4
nvidia

Description: The NVIDIA Qwen3.6-27B NVFP4 model is the quantized version of Alibaba's Qwen3.6-27B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.6-27B NVFP4 model is quan…

Open Source 27.0B ★ 38.0 ↓ 614.8K
MiMo-V2.5 logo
MiMo-V2.5
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio  

Open Source ★ 38.0 ↓ 221.8K
MiMo-V2.6-Flash-RL logo
MiMo-V2.6-Flash-RL
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio   💻 Xiaomi MiMo Desktop  

Open Source ★ 38.0 ↓ 55.6K
MiMo-V2.6-Flash-MOPD logo
MiMo-V2.6-Flash-MOPD
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨Xiaomi MiMo Studio   💻 Xiaomi MiMo Desktop  

Open Source ★ 38.0 ↓ 6.6K
MiMo-V2.5-Base logo
MiMo-V2.5-Base
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio  

Open Source ★ 38.0 ↓ 677
Qwen3.8-27B logo
Qwen3.8-27B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, TokenSpeed, etc.

Open Source 27.0B ★ 34.0 ↓ 6.8M
Qwen3.8-27B-FP8 logo
Qwen3.8-27B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, TokenSpeed, etc. The quantization method is fine-g…

Open Source 27.0B ★ 34.0 ↓ 4.7M
Qwen3.8-27B-NVFP4 logo
Qwen3.8-27B-NVFP4
nvidia

Description: The NVIDIA Qwen3.8-27B NVFP4 model is a quantized version of Alibaba's Qwen3.8-27B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information on the model, please check here. The model is quantized with Mod…

Open Source 27.0B ★ 34.0 ↓ 622.9K
Qwen3.5-122B-A10B-FP8 logo
Qwen3.5-122B-A10B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

Multimodal 122.0B ★ 33.0 ↓ 1.7M