AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

63 models for "Open Weights" Compare
OpenAI: GPT-6 Astra Pro logo
OpenAI: GPT-6 Astra Pro
openai

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

Closed Source ★ 53.0
Meta: Muse Spark 1.3 Contributor logo
Meta: Muse Spark 1.3 Contributor
meta

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...

Closed Source ★ 48.0
SpaceXAI: Grok 4.7 logo
SpaceXAI: Grok 4.7
x-ai

Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...

Closed Source ★ 46.0
Z.ai: GLM 5.3 (batch) logo
Z.ai: GLM 5.3 (batch)
z-ai

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Open Source 744.0B ★ 45.0
DeepSeek: DeepSeek V4.1 Flash (batch) logo
DeepSeek: DeepSeek V4.1 Flash (batch)
deepseek

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

Multimodal 552.0B ★ 39.0
DeepSeek: DeepSeek Flash Latest logo
DeepSeek: DeepSeek Flash Latest
~deepseek

This model always redirects to the latest model in the DeepSeek Flash family.

Multimodal 552.0B ★ 39.0
🤖
Xiaomi: MiMo-V2.6-Flash
xiaomi

MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...

Multimodal 309.0B ★ 38.0
inclusionAI: Ling 3.0 Flash Fin (free) logo
inclusionAI: Ling 3.0 Flash Fin (free)
inclusionai

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

Open Source 124.0B ★ 38.0
OpenAI: GPT-6 Luna logo
OpenAI: GPT-6 Luna
openai

GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

Closed Source ★ 37.0
DeepSeek: DeepSeek Pro Latest logo
DeepSeek: DeepSeek Pro Latest
~deepseek

This model always redirects to the latest model in the DeepSeek Pro family.

Open Source 1600.0B ★ 36.0
Ling-3.0-flash-VL-int4 logo
Ling-3.0-flash-VL-int4
inclusionAI

🤗 Hugging Face      🤖 ModelScope   

Multimodal 124.0B ★ 25.0 ↓ 1.4K
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Base-BF16 logo
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Base-BF16
nvidia

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Base-BF16

Open Source 30.0B ★ 24.0 ↓ 143.2K