LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and...
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
This model always redirects to the latest model in the OpenAI GPT family.
This model always redirects to the latest model in the Anthropic Claude Sonnet family.
This model always redirects to the latest model in the Google Gemini Flash family.
This model always redirects to the latest model in the MoonshotAI Kimi family.
This model always redirects to the latest model in the Google Gemini Pro family.
This model always redirects to the latest model in the OpenAI GPT Mini family.
This model always redirects to the latest model in the Anthropic Claude Haiku family.
NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...