AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

63 models for "Open Weights" Compare
North-Mini-Code-1.0-w4a16 logo
North-Mini-Code-1.0-w4a16
CohereLabs

North Mini Code is an open weights research release of a 30B-A3B parameter model optimized for code generation, agentic software engineering, and terminal tasks.

Code ★ 10.0 ↓ 979
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 1.5M
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 905.2K
Qwen3-Coder-Next logo
Qwen3-Coder-Next
Qwen

Today, we're announcing Qwen3-Coder-Next , an open-weight language model designed specifically for coding agents and local development. It features the following key enhancements:

Code 80.0B ★ 9.0 ↓ 599.2K
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-FP8
nvidia

The post-training data has a cutoff date of November 28, 2025\. The pre-training data has a cutoff date of June 25, 2025\.

Open Source 30.0B ★ 9.0 ↓ 246K
NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16 logo
NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16
nvidia

NVIDIA-Nemotron-3-Nano-30B-A3B-Base-BF16

Open Source 30.0B ★ 9.0 ↓ 79.2K
Mistral: Mistral Large 3 2512 (batch) logo
Mistral: Mistral Large 3 2512 (batch)
mistralai

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Open Source 675.0B ★ 9.0
tiny-aya-global logo
tiny-aya-global
CohereLabs

Open Source 3.35B ★ 5.0 ↓ 8.3K
gemma-3-1b-it logo
gemma-3-1b-it
google

Open Source 1.0B ↓ 3.1M
gemma-2-9b-it logo
gemma-2-9b-it
google

Open Source 9.0B ↓ 968.5K
SmolLM3-3B logo
SmolLM3-3B
HuggingFaceTB

1. Model Summary 2. How to use 3. Evaluation 4. Training 5. Limitations 6. License

Open Source 3.0B ↓ 640.4K
Qwen3-4B-Thinking-2507 logo
Qwen3-4B-Thinking-2507
Qwen

Over the past three months, we have continued to scale the thinking capability of Qwen3-4B, improving both the quality and depth of reasoning. We are pleased to introduce Qwen3-4B-Thinking-2507 , featuring the following key enhancements:

Reasoning 4.0B ↓ 603.7K