LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Baidu: ERNIE 4.5 VL 424B A47B
baidu
ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token. It is trained jointly on text and image data...
Multimodal
Qwen: Qwen3 235B A22B Thinking 2507
qwen
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Closed Source
MoonshotAI: Kimi K2 0905
moonshotai
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
Closed Source