AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,315 models for "Transformer" Compare
🤖
Meta-Llama-3.1-70B-Instruct
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 70.0B ↓ 30.4K
🤖
Baichuan2-7B-Chat
baichuan-inc

🦉GitHub 💬WeChat 百川API支持搜索增强和192K长窗口,新增百川搜索增强知识库、限时免费! 🚀 百川大模型在线对话平台 已正式向公众开放 🎉

Open Source 7.0B ↓ 29.2K
🤖
InternVL3_5-30B-A3B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 30.85B ↓ 28.5K
🤖
Yi-34B-Chat
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 28.3K
🤖
ERNIE-4.5-21B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 27.4K
🤖
MiniCPM-SALA
openbmb

GitHub Repo Technical Report Join Us 👋 Contact us in Discord and WeChat

Open Source 9.0B ↓ 27K
🤖
truthfulqa-truth-judge-llama2-7B
allenai

This model is built based on LLaMa2 7B in replacement of the truthfulness/informativeness judge models that were originally introduced in the TruthfulQA paper. That model is based on OpenAI's Curie engine using their finetuning API. However, as of February 08, 2024, OpenAI has ta…

Open Source 7.0B ↓ 26.7K
🤖
MiniCPM3-4B
openbmb

MiniCPM Repo MiniCPM Paper MiniCPM-V Repo Join us in Discord and WeChat

Open Source 4.0B ↓ 26.6K
🤖
LFM2-VL-1.6B
LiquidAI

LFM2‑VL is Liquid AI's first series of multimodal models, designed to process text and images with variable resolutions. Built on the LFM2 backbone, it is optimized for low-latency and edge AI applications.

Multimodal 1.6B ↓ 25.7K
🤖
truthfulqa-info-judge-llama2-7B
allenai

This model is built based on LLaMa2 7B in replacement of the truthfulness/informativeness judge models that were originally introduced in the TruthfulQA paper. That model is based on OpenAI's Curie engine using their finetuning API. However, as of February 08, 2024, OpenAI has ta…

Open Source 7.0B ↓ 24.5K
🤖
deepseek-llm-7b-chat
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 7.0B ↓ 24.3K
🤖
InternVL3_5-38B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.4B ↓ 24.1K