AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
ERNIE-4.5-300B-A47B-Base-Paddle logo
ERNIE-4.5-300B-A47B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 177
ERNIE-4.5-300B-A47B-W4A8C8-TP4-Paddle logo
ERNIE-4.5-300B-A47B-W4A8C8-TP4-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 144
ERNIE-4.5-300B-A47B-FP8-Paddle logo
ERNIE-4.5-300B-A47B-FP8-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 300.0B ★ 8.0 ↓ 139
ERNIE-4.5-300B-A47B-2Bits-TP2-Paddle logo
ERNIE-4.5-300B-A47B-2Bits-TP2-Paddle
baidu

The advanced capabilities of the ERNIE 4.5 models, particularly the MoE-based A47B and A3B series, are underpinned by several key technical innovations:

Open Source 300.0B ★ 8.0 ↓ 128
ERNIE-4.5-300B-A47B-2Bits-TP4-Paddle logo
ERNIE-4.5-300B-A47B-2Bits-TP4-Paddle
baidu

The advanced capabilities of the ERNIE 4.5 models, particularly the MoE-based A47B and A3B series, are underpinned by several key technical innovations:

Open Source 300.0B ★ 8.0 ↓ 85
NVIDIA-Nemotron-3-Nano-4B-BF16 logo
NVIDIA-Nemotron-3-Nano-4B-BF16
nvidia

The pretraining data has a cutoff date of September 2024\.

Open Source 4.0B ★ 7.0 ↓ 1.9M
Olmo-3.1-32B-Think logo
Olmo-3.1-32B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 7.0 ↓ 19.2K
Olmo-3-7B-Think logo
Olmo-3-7B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 6.0 ↓ 489.1K
Olmo-3.1-32B-Instruct logo
Olmo-3.1-32B-Instruct
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 6.0 ↓ 18.7K
Olmo-3-7B-Think-DPO logo
Olmo-3-7B-Think-DPO
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 6.0 ↓ 8.7K
Olmo-3-7B-Think-SFT logo
Olmo-3-7B-Think-SFT
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 6.0 ↓ 8.5K
Olmo-3.1-32B-Instruct-DPO logo
Olmo-3.1-32B-Instruct-DPO
allenai

Model Card for Olmo-3.1-32B-Instruct-DPO

Open Source 32.0B ★ 6.0 ↓ 2.5K