AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
RedPajama-INCITE-Base-3B-v1 logo
RedPajama-INCITE-Base-3B-v1
togethercomputer

RedPajama-INCITE-Base-3B-v1 was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research re…

Open Source 3.0B ↓ 4.8K
internlm2-chat-7b-sft logo
internlm2-chat-7b-sft
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 4.4K
internlm2_5-7b logo
internlm2_5-7b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 4.4K
bloom-7b1 logo
bloom-7b1
bigscience

BLOOM LM BigScience Large Open-science Open-access Multilingual Language Model Model Card

Open Source 7.07B ↓ 4.4K
InternVL3-9B logo
InternVL3-9B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 9.1B ↓ 4.3K
granite-3.3-8b-instruct-FP8 logo
granite-3.3-8b-instruct-FP8
ibm-granite

[!NOTE] This repository contains the FP8 version of Granite-3.3-8b-Instruct . Please reference the base model's full model card here: https://huggingface.co/ibm-granite/granite-3.3-8b-instruct

Open Source 8.0B ↓ 4.3K
granite-3.1-1b-a400m-base logo
granite-3.1-1b-a400m-base
ibm-granite

Model Summary: Granite-3.1-1B-A400M-Base extends the context length of Granite-3.0-1B-A400M-Base from 4K to 128K using a progressive training strategy by increasing the supported context length in increments while adjusting RoPE theta until the model has successfully adapted to d…

Open Source 1.0B ↓ 4.2K
InternVL3-1B-Instruct logo
InternVL3-1B-Instruct
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 1.0B ↓ 4K
Olmo-Hybrid-Instruct-SFT-7B logo
Olmo-Hybrid-Instruct-SFT-7B
allenai

We expand on our Olmo model series by introducing Olmo Hybrid, a new 7B hybrid RNN model in the Olmo family. Olmo Hybrid dramatically outperforms Olmo 3 in final performance, consistently showing roughly 2x data efficiency on core evals over the course of our pretraining run. We…

Open Source 7.0B ↓ 3.7K
GLM-4.5-FP8 logo
GLM-4.5-FP8
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.5 technical blog . 📍 Use GLM-4.5 API services on Z.ai API Platform (Global) or Zhipu AI Open Platform (Mainland China) . 👉 One click to GLM-4.5 .

Open Source ↓ 3.7K
OLMo-2-1124-7B-DPO logo
OLMo-2-1124-7B-DPO
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 3.7K
Hunyuan-1.8B-Instruct logo
Hunyuan-1.8B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 1.8B ↓ 3.6K