AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
deepseekcoder-33b-codeqwen-align-subset logo
deepseekcoder-33b-codeqwen-align-subset
bigcode

This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.

Code 33.0B ↓ 170
ERNIE-4.5-0.3B-Base-Paddle logo
ERNIE-4.5-0.3B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 0.3B ↓ 167
Gemma-SEA-LION-v4-27B-VL logo
Gemma-SEA-LION-v4-27B-VL
aisingapore

SEA-LION-VL is an instruct-tuned vision-text model for the Southeast Asia (SEA) region.

Multimodal 27.0B ↓ 163
ERNIE-4.5-VL-28B-A3B-Base-Paddle logo
ERNIE-4.5-VL-28B-A3B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 28.0B ↓ 160
ERNIE-4.5-VL-424B-A47B-Paddle logo
ERNIE-4.5-VL-424B-A47B-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 159
ERNIE-4.5-VL-424B-A47B-Base-Paddle logo
ERNIE-4.5-VL-424B-A47B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 147
ERNIE-4.5-21B-A3B-Base-Paddle logo
ERNIE-4.5-21B-A3B-Base-Paddle
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 140
diafill-sarashina2.2-3b-instruct logo
diafill-sarashina2.2-3b-instruct
sbintuitions

Model Summary DiaFill is a Japanese dialogue script generation model designed to produce natural, spoken-style dialogue scripts rich in fillers and brief utternaces. Unlike typical assistant models that respond to users, this model is fine-tuned to generate a multi-turn dialogue…

Open Source 3.0B ↓ 136
llama-3-youko-70b-instruct logo
llama-3-youko-70b-instruct
rinna

Llama 3 Youko 70B Instruct (rinna/llama-3-youko-70b-instruct)

Open Source 70.0B ↓ 130
bloom-3b-intermediate logo
bloom-3b-intermediate
bigscience

WARNING: The checkpoints on this repo are not fully trained model. Evaluations of intermediary checkpoints and the final model will be added when conducted (see below).

Open Source 3.0B ↓ 127
diafill-llm-jp-3.1-13b-instruct4 logo
diafill-llm-jp-3.1-13b-instruct4
sbintuitions

Model Summary DiaFill is a Japanese dialogue script generation model designed to produce natural, spoken-style dialogue scripts rich in fillers and brief utternaces. Unlike typical assistant models that respond to users, this model is fine-tuned to generate a multi-turn dialogue…

Open Source 13.0B ↓ 124
bloom-1b7-intermediate logo
bloom-1b7-intermediate
bigscience

WARNING: The checkpoints on this repo are not fully trained model. Evaluations of intermediary checkpoints and the final model will be added when conducted (see below).

Open Source ↓ 120