AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,116 models for "Chat" Compare
🤖
cogvlm2-llama3-chat-19B
zai-org

👋 Wechat · 💡 Online Demo · 🎈 Github Page · 📑 Paper 📍Experience the larger-scale CogVLM model on the ZhipuAI Open Platform .

Multimodal 19.0B ↓ 8.5K
🤖
Infinity-Instruct-3M-0625-Llama3-8B
BAAI

Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][🤗] (would be released soon)

Open Source 8.0B ↓ 8.5K
🤖
tulu-2-7b
allenai

Tulu is a series of language models that are trained to act as helpful assistants. Tulu 2 7B is a fine-tuned version of Llama 2 that was trained on a mix of publicly available, synthetic and human datasets.

Open Source 7.0B ↓ 8.5K
🤖
granite-3b-code-instruct-2k
ibm-granite

⚠️ DEPRECATION WARNING ⚠️ ⚠️ NOT RECOMMENDED FOR USE IN NEW PROJECTS ⚠️

Code 3.0B ↓ 8.5K
🤖
OLMo-2-1124-13B-Instruct
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 13.0B ↓ 8.3K
🤖
granite-3.2-2b-instruct
ibm-granite

Model Summary: Granite-3.2-2B-Instruct is an 2-billion-parameter, long-context AI model fine-tuned for thinking capabilities. Built on top of Granite-3.1-2B-Instruct, it has been trained using a mix of permissively licensed open-source datasets and internally generated synthetic…

Open Source 2.0B ↓ 8.3K
🤖
Llama-2-70b-chat-hf
meta-llama

Open Source 70.0B ↓ 8.3K
🤖
DeepSeek-R1-Zero
deepseek-ai

We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With R…

Reasoning ↓ 8.2K
🤖
SmolLM-1.7B-Instruct
HuggingFaceTB

SmolLM is a series of small language models available in three sizes: 135M, 360M, and 1.7B parameters.

Open Source 1.7B ↓ 8.2K
🤖
Ring-2.5-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 Experience Link Coming Soon~

Open Source ↓ 8.2K
🤖
InternVL-Chat-V1-2
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal ↓ 8.2K
🤖
granite-3.0-2b-instruct
ibm-granite

Model Summary: Granite-3.0-2B-Instruct is a 2B parameter model finetuned from Granite-3.0-2B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. This model is developed using a diverse set of techniques…

Open Source 2.0B ↓ 8.2K