AI Agent Hub

LLM Models · Reasoning

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

Compare
🤖
DeepSeek-R1
DeepSeek-R1

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

Reasoning ↓ 2M
🤖
NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4

Reasoning ↓ 1.1M
🤖
NVIDIA-Nemotron-3-Super-120B-A12B-BF16
NVIDIA-Nemotron-3-Super-120B-A12B-BF16

Reasoning ↓ 1M
🤖
DeepSeek-R1-0528-Qwen3-8B
DeepSeek-R1-0528-Qwen3-8B

Reasoning ↓ 1M
🤖
DeepSeek-R1-Distill-Qwen-32B
DeepSeek-R1-Distill-Qwen-32B

Reasoning ↓ 584.4K
🤖
DeepSeek-R1-Distill-Qwen-14B
DeepSeek-R1-Distill-Qwen-14B

Reasoning ↓ 486.8K
🤖
DeepSeek-R1-Distill-Qwen-1.5B
DeepSeek-R1-Distill-Qwen-1.5B

Reasoning ↓ 486.6K
🤖
DeepSeek-R1-Distill-Llama-8B
DeepSeek-R1-Distill-Llama-8B

Reasoning ↓ 389.2K
🤖
DeepSeek-R1-Distill-Qwen-7B
DeepSeek-R1-Distill-Qwen-7B

Reasoning ↓ 369.5K
🤖
DeepSeek-R1-0528
DeepSeek-R1-0528

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

Reasoning ↓ 198.3K
🤖
DeepSeek-R1-Distill-Llama-70B
DeepSeek-R1-Distill-Llama-70B

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

Reasoning ↓ 93.7K
🤖
Amazon: Nova Premier 1.0
amazon

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

Reasoning