LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Qwen2.5-32B-Instruct
Qwen2.5-32B-Instruct
Open Source
↓ 2M
Qwen2.5-Coder-14B-Instruct
Qwen2.5-Coder-14B-Instruct
Code
↓ 2M
DeepSeek-R1
DeepSeek-R1
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
Reasoning
↓ 2M
Gemma-4-31B-IT-NVFP4
Gemma-4-31B-IT-NVFP4
Open Source
↓ 2M
Ternary-Bonsai-27B-mlx-2bit
Ternary-Bonsai-27B-mlx-2bit
Open Source
↓ 2M
Qwen3.6-27B-AWQ-INT4
Qwen3.6-27B-AWQ-INT4
Open Source
↓ 2M
Bonsai-27B-mlx-1bit
Bonsai-27B-mlx-1bit
Open Source
↓ 2M
GLM-4.7-Flash
GLM-4.7-Flash
Open Source
↓ 1.9M
TinyLlama-1.1B-Chat-v1.0
TinyLlama-1.1B-Chat-v1.0
Open Source
↓ 1.9M
Qwen2.5-Coder-14B-Instruct-AWQ
Qwen2.5-Coder-14B-Instruct-AWQ
Code
↓ 1.9M
Ornith-1.0-9B
Ornith-1.0-9B
Open Source
↓ 1.8M
Qwen3-14B
Qwen3-14B
Open Source
↓ 1.8M