AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Llama-xLAM-2-8b-fc-r logo
Llama-xLAM-2-8b-fc-r
Salesforce

Large Action Models (LAMs) are advanced language models designed to enhance decision-making by translating user intentions into executable actions. As the brains of AI agents , LAMs autonomously plan and execute tasks to achieve specific goals, making them invaluable for automati…

Open Source 8.0B ↓ 95.3K
paligemma-3b-mix-224 logo
paligemma-3b-mix-224
google

Multimodal 2.92B ↓ 91.8K
granite-3.3-8b-instruct logo
granite-3.3-8b-instruct
ibm-granite

Model Summary: Granite-3.3-8B-Instruct is a 8-billion parameter 128K context length language model fine-tuned for improved reasoning and instruction-following capabilities. Built on top of Granite-3.3-8B-Base, the model delivers significant gains on benchmarks for measuring gener…

Open Source 8.0B ↓ 81.8K
opt-1.3b logo
opt-1.3b
facebook

OPT : Open Pre-trained Transformer Language Models

Open Source 1.3B ↓ 80.2K
DeepSeek-R1-Distill-Llama-70B logo
DeepSeek-R1-Distill-Llama-70B
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning 70.0B ↓ 61.9K
Meta-Llama-3-70B-Instruct logo
Meta-Llama-3-70B-Instruct
meta-llama

Open Source 70.0B ↓ 59.5K
granite-4.0-3b-vision logo
granite-4.0-3b-vision
ibm-granite

Model Summary: Granite-4.0-3B-Vision is a vision-language model (VLM) designed for enterprise-grade document data extraction. It focuses on specialized, complex extraction tasks that ultracompact models often struggle with:

Multimodal 4.0B ↓ 56.3K
Phi-4-mini-reasoning logo
Phi-4-mini-reasoning
microsoft

Phi-4-mini-reasoning is a lightweight open model built upon synthetic data with a focus on high-quality, reasoning dense data further finetuned for more advanced math reasoning capabilities. The model belongs to the Phi-4 model family and supports 128K token context length.

Open Source ↓ 56.2K
granite-guardian-4.1-8b logo
granite-guardian-4.1-8b
ibm-granite

Granite Guardian 4.1 8B introduces improved Bring Your Own Criteria (BYOC) support, enabling users to define arbitrary judging criteria beyond the pre-baked safety and hallucination detectors. The model can now faithfully evaluate complex, multi-part requirements such as formatti…

Open Source 8.0B ↓ 55.4K
granite-4.0-tiny-preview logo
granite-4.0-tiny-preview
ibm-granite

Model Summary: Granite-4-Tiny-Preview is a 7B parameter fine-grained hybrid mixture-of-experts (MoE) instruct model fine-tuned from Granite-4.0-Tiny-Base-Preview using a combination of open source instruction datasets with permissive license and internally collected synthetic dat…

Open Source 7.0B ↓ 54.1K
OLMo-1B-hf logo
OLMo-1B-hf
allenai

OLMo is a series of O pen L anguage Mo dels designed to enable the science of language models. The OLMo models are trained on the Dolma dataset. We release all code, checkpoints, logs (coming soon), and details involved in training these models. This model has been converted from…

Code 1.0B ↓ 53.4K
granite-3.2-8b-instruct logo
granite-3.2-8b-instruct
ibm-granite

Model Summary: Granite-3.2-8B-Instruct is an 8-billion-parameter, long-context AI model fine-tuned for thinking capabilities. Built on top of Granite-3.1-8B-Instruct, it has been trained using a mix of permissively licensed open-source datasets and internally generated synthetic…

Open Source 8.17B ↓ 52.9K