AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Olmo-3-7B-Instruct logo
Olmo-3-7B-Instruct
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 5.0 ↓ 473.2K
Olmo-3-7B-Instruct-SFT logo
Olmo-3-7B-Instruct-SFT
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 5.0 ↓ 36.5K
Olmo-3-7B-Instruct-DPO logo
Olmo-3-7B-Instruct-DPO
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 7.0B ★ 5.0 ↓ 14.7K
Ministral-3-3B-Instruct-2512-ONNX logo
Ministral-3-3B-Instruct-2512-ONNX
mistralai

[!Tip] This model was contributed by Xenova from Hugging Face. We sincerely appreciate the integration and community collaboration. While preliminary functionality checks have been performed, comprehensive testing has not yet been completed. We recommend you to proceed with cauti…

Open Source 3.0B ★ 5.0 ↓ 507
OLMo-7B logo
OLMo-7B
allenai

For transformers versions v4.40.0 or newer, we suggest using OLMo 7B HF instead.

Open Source 7.0B ★ 3.0 ↓ 4.9K
opt-125m logo
opt-125m
facebook

OPT : Open Pre-trained Transformer Language Models

Open Source ↓ 6.3M
Llama-3.1-8B-Instruct logo
Llama-3.1-8B-Instruct
meta-llama

Open Source 8.0B ↓ 5.7M
Florence-2-base logo
Florence-2-base
microsoft

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Multimodal 0.23B ↓ 3.1M
Mistral-7B-Instruct-v0.2 logo
Mistral-7B-Instruct-v0.2
mistralai

py from mistral common.tokens.tokenizers.mistral import MistralTokenizer from mistral common.protocol.instruct.messages import UserMessage from mistral common.protocol.instruct.request import ChatCompletionRequest

Open Source 7.0B ↓ 1.6M
DeepSeek-R1 logo
DeepSeek-R1
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning ↓ 1.1M
DeepSeek-R1-Distill-Qwen-1.5B logo
DeepSeek-R1-Distill-Qwen-1.5B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 1.5B ↓ 1M
MiniMax-M2.5-NVFP4 logo
MiniMax-M2.5-NVFP4
nvidia

Description: The NVIDIA MiniMax-M2.5-NVFP4 model is the quantized version of MiniMax's MiniMax-M2.5 model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA MiniMax-M2.5 NVFP4 model is q…

Open Source ↓ 771.3K