AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

377 models for "MoE" Compare
🤖
Qwen3-VL-8B-Instruct-FP8
Qwen

This repository contains an FP8 quantized version of the Qwen3-VL-8B-Instruct model. The quantization method is fine-grained fp8 quantization with block size of 128, and its performance metrics are nearly identical to those of the original BF16 model. Enjoy!

Multimodal 8.0B ↓ 2.3M
🤖
GLM-4.7-Flash
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.7 technical blog , technical report(GLM-4.5) . 📍 Use GLM-4.7-Flash API services on Z.ai API Platform. 👉 One click to GLM-4.7 .

Open Source ↓ 1.9M
🤖
Qwen3-1.7B-Base
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Building upon extensive advancements in training data, model architecture, and optimization techniques, Qwen3 delivers the followin…

Open Source 1.7B ↓ 1.9M
🤖
DeepSeek-R1
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning ↓ 1.9M
🤖
Qwen3-4B-Base
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Building upon extensive advancements in training data, model architecture, and optimization techniques, Qwen3 delivers the followin…

Open Source 4.0B ↓ 1.9M
🤖
Qwen3-14B
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 14.8B ↓ 1.8M
🤖
Qwen3-8B-AWQ
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities,…

Open Source 8.0B ↓ 1.8M
🤖
Qwen3-30B-A3B-Instruct-2507
Qwen

We introduce the updated version of the Qwen3-30B-A3B non-thinking mode , named Qwen3-30B-A3B-Instruct-2507 , featuring the following key enhancements:

Open Source 30.0B ↓ 1.4M
🤖
Qwen3-Coder-30B-A3B-Instruct-FP8
Qwen

Qwen3-Coder is available in multiple sizes. Today, we're excited to introduce Qwen3-Coder-30B-A3B-Instruct-FP8 . This streamlined model maintains impressive performance and efficiency, featuring the following key enhancements:

Code 30.5B ↓ 1.1M
🤖
DeepSeek-V3
deepseek-ai

We present DeepSeek-V3, a strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. To achieve efficient inference and cost-effective training, DeepSeek-V3 adopts Multi-head Latent Attention (MLA) and DeepSeekMoE architectures, w…

Open Source ↓ 1.1M
🤖
DeepSeek-OCR-2
deepseek-ai

🌟 Github 📥 Model Download 📄 Paper Link 📄 Arxiv Paper Link DeepSeek-OCR 2: Visual Causal Flow Explore more human-like visual encoding.

Multimodal 3.0B ↓ 1.1M
🤖
Qwen3-0.6B-Base
Qwen

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Building upon extensive advancements in training data, model architecture, and optimization techniques, Qwen3 delivers the followin…

Open Source 0.6B ↓ 989.1K