AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Yi-6B-200K logo
Yi-6B-200K
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 22.9K
deepseek-llm-7b-chat logo
deepseek-llm-7b-chat
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 7.0B ↓ 22.2K
ERNIE-4.5-21B-A3B-PT logo
ERNIE-4.5-21B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 19.1K
InternVL3-2B logo
InternVL3-2B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 2.1B ↓ 18.7K
Yi-1.5-9B-Chat logo
Yi-1.5-9B-Chat
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 17.9K
Llama-2-7b-chat-hf logo
Llama-2-7b-chat-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 7B fine-tuned model, optimized for dialogue use cases and converted for the Hugging Face Transformers forma…

Open Source 7.0B ↓ 17.9K
InternVL3-38B logo
InternVL3-38B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.4B ↓ 16.8K
starcoder logo
starcoder
bigcode

Code 15.5B ↓ 16.6K
OLMo-7B-0724-hf logo
OLMo-7B-0724-hf
allenai

OLMo is a series of O pen L anguage Mo dels designed to enable the science of language models. The OLMo models are trained on the Dolma dataset. We release all code, checkpoints, logs, and details involved in training these models.

Open Source 7.0B ↓ 16.1K
Yi-1.5-34B-Chat logo
Yi-1.5-34B-Chat
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 34.0B ↓ 15.6K
ERNIE-4.5-0.3B-PT logo
ERNIE-4.5-0.3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 0.36B ↓ 14.9K
Yi-1.5-9B logo
Yi-1.5-9B
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 14.8K