AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Qwen2.5-0.5B-DuDi logo
Qwen2.5-0.5B-DuDi
aisingapore

This model is a fine-tuned version of Qwen2.5-0.5B. It has been trained using TRL.

Open Source 0.5B ↓ 24
Qwen3-0.6B-Base-DuDi logo
Qwen3-0.6B-Base-DuDi
aisingapore

This model is a fine-tuned version of Qwen3-0.6B-Base. It has been trained using TRL.

Open Source 0.6B ↓ 17
Qwen2.5-1.5B-DuDi logo
Qwen2.5-1.5B-DuDi
aisingapore

This model is a fine-tuned version of Qwen2.5-1.5B. It has been trained using TRL.

Open Source 1.5B ↓ 16
Llama3.2-1B-DuDi logo
Llama3.2-1B-DuDi
aisingapore

This model is a fine-tuned version of Llama-3.2-1B. It has been trained using TRL.

Open Source 1.0B ↓ 16
Meta: Llama Guard 4 12B logo
Meta: Llama Guard 4 12B
meta-llama

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

Multimodal 12.0B
🤖
Inference.net: Schematron V2 Small
inference-net

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

Open Source 3.0B
NVIDIA: Nemotron 3.5 Content Safety logo
NVIDIA: Nemotron 3.5 Content Safety
nvidia

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

Multimodal 4.0B
Mistral: Mixtral 8x22B Instruct logo
Mistral: Mixtral 8x22B Instruct
mistralai

Mistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size. Its strengths include: - strong math, coding,...

Closed Source
Magnum v4 72B logo
Magnum v4 72B
anthracite-org

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-

Closed Source
AionLabs: Aion-RP 1.0 (8B) logo
AionLabs: Aion-RP 1.0 (8B)
aion-labs

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

Closed Source
TheDrummer: Skyfall 36B V2 logo
TheDrummer: Skyfall 36B V2
thedrummer

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

Closed Source
Venice: Uncensored logo
Venice: Uncensored
cognitivecomputations

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

Closed Source