AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Yi-34B logo
Yi-34B
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 10.2K
OLMo-2-0325-32B-Instruct logo
OLMo-2-0325-32B-Instruct
allenai

OLMo 2 32B Instruct March 2025 is post-trained variant of the OLMo-2 32B March 2025 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training on this dataset. Tülu 3 is designe…

Open Source 32.0B ↓ 9.9K
Yi-1.5-34B logo
Yi-1.5-34B
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 34.0B ↓ 9.6K
Yi-34B-200K logo
Yi-34B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 9.2K
Yi-1.5-9B-Chat-16K logo
Yi-1.5-9B-Chat-16K
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 9K
Yi-9B logo
Yi-9B
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 9.0B ↓ 9K
Yi-1.5-34B-Chat-16K logo
Yi-1.5-34B-Chat-16K
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 34.0B ↓ 8.9K
Yi-9B-200K logo
Yi-9B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 9.0B ↓ 8.9K
bloom-1b1 logo
bloom-1b1
bigscience

BLOOM LM BigScience Large Open-science Open-access Multilingual Language Model Model Card

Open Source 1.065B ↓ 8.9K
Yi-1.5-34B-32K logo
Yi-1.5-34B-32K
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 34.0B ↓ 8.9K
Yi-1.5-9B-32K logo
Yi-1.5-9B-32K
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 8.8K
DeepSeek-R1-Zero logo
DeepSeek-R1-Zero
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning ↓ 8.7K