AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
aya-23-8B logo
aya-23-8B
CohereLabs

Open Source 8.0B ↓ 8.5K
tulu-2-7b logo
tulu-2-7b
allenai

Tulu is a series of language models that are trained to act as helpful assistants. Tulu 2 7B is a fine-tuned version of Llama 2 that was trained on a mix of publicly available, synthetic and human datasets.

Open Source 7.0B ↓ 8.5K
Yi-1.5-6B-Chat logo
Yi-1.5-6B-Chat
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 6.0B ↓ 8.3K
falcon-mamba-7b-instruct logo
falcon-mamba-7b-instruct
tiiuae

Model card for FalconMamba Instruct model

Open Source 7.0B ↓ 8.2K
InternVL3-78B logo
InternVL3-78B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 78.4B ↓ 8.1K
Mini-InternVL2-4B-DA-Medical logo
Mini-InternVL2-4B-DA-Medical
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[🆕 Blog\]](https://internvl.github.io/blog/) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 1.0\]](https://arxiv.org/abs/2312.14238) [\[📜 InternVL 1.5\]](https://arxiv.org/abs/2404.16821) [\[📜 InternVL…

Multimodal 4.0B ↓ 7.8K
Jamba-v0.1 logo
Jamba-v0.1
ai21labs

This is the base version of the Jamba model. We’ve since released a better, instruct-tuned version, Jamba-1.5-Mini. For even greater performance, check out the scaled-up Jamba-1.5-Large.

Open Source ↓ 7.6K
granite-guardian-3.2-3b-a800m logo
granite-guardian-3.2-3b-a800m
ibm-granite

Granite Guardian 3.2 3B-A800M is a fine-tuned Granite 3.2 3B-A800M instruct model designed to detect risks in prompts and responses. It can help with risk detection along many key dimensions catalogued in the IBM AI Risk Atlas. It is trained on unique data comprising human annota…

Open Source 3.3B ↓ 7.3K
xLAM-1b-fc-r logo
xLAM-1b-fc-r
Salesforce

[Homepage] [APIGen Paper] [ActionStudio Paper] [Discord] [Dataset] [Github]

Open Source 1.35B ↓ 6.5K
InternVL3-14B logo
InternVL3-14B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 15.1B ↓ 6.3K
Llama-3.1-Tulu-3-8B logo
Llama-3.1-Tulu-3-8B
allenai

Tülu 3 is a leading instruction following model family, offering a post-training package with fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern techniques. This is one step of a bigger process to training fully open-source models, lik…

Open Source 8.0B ↓ 6.3K
OLMo-2-1124-7B-SFT logo
OLMo-2-1124-7B-SFT
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 6.2K