AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

Compare
🤖
Upstage: Solar Pro 4
upstage

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive...

Closed Source
🤖
Sakana: Sakana Namazu
sakana

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...

Closed Source
🤖
NVIDIA: Nemotron 3.5 Lightning (free)
nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Closed Source
🤖
NVIDIA: Nemotron 3.5 Lightning
nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...

Closed Source
🤖
LiquidAI: LFM2.5-2.6B (free)
liquid

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

Closed Source
🤖
SpaceXAI: Grok 4.6
x-ai

Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

Closed Source
🤖
DeepSeek: DeepSeek V4 Pro 0813 (batch)
deepseek

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

Closed Source
🤖
DeepSeek: DeepSeek V4 Pro 0813
deepseek

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

Closed Source
🤖
ByteDance Seed: Seed-2.0-Code
bytedance-seed

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...

Code
🤖
Qwen: Qwen3.8 2.4T A95B (batch)
qwen

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Closed Source
🤖
Qwen: Qwen3.8 2.4T A95B
qwen

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Closed Source
🤖
ByteDance Seed: Seed 2.1 Turbo
bytedance-seed

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

Closed Source