AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
granite-swash-3b-a600m logo
granite-swash-3b-a600m
ibm-granite

Granite-SWASH-3B-a600M (Sliding Window Attention + Sinks Hybrid)

Open Source 3.0B ↓ 18.1K
Yi-1.5-9B-Chat logo
Yi-1.5-9B-Chat
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 17.9K
Molmo2-ER logo
Molmo2-ER
allenai

Molmo2-ER (Embodied Reasoning) is a 4B vision–language model specialized for the embodied perception skills that downstream action models depend on: scene understanding, pixel-accurate pointing, multi-image and egocentric–exocentric correspondence, and video temporal reasoning.

Multimodal 4.0B ↓ 17.8K
granite-3.0-1b-a400m-base logo
granite-3.0-1b-a400m-base
ibm-granite

Model Summary: Granite-3.0-1B-A400M-Base is a decoder-only language model to support a variety of text-to-text generation tasks. It is trained from scratch following a two-stage training strategy. In the first stage, it is trained on 8 trillion tokens sourced from diverse domains…

Open Source 1.3B ↓ 17.6K
internlm2-base-20b logo
internlm2-base-20b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 20.0B ↓ 17.3K
InternVL3-38B logo
InternVL3-38B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.4B ↓ 16.8K
GLM-4-9B-0414 logo
GLM-4-9B-0414
zai-org

The GLM family welcomes new members, the GLM-4-32B-0414 series models, featuring 32 billion parameters. Its performance is comparable to OpenAI’s GPT series and DeepSeek’s V3/R1 series. It also supports very user-friendly local deployment features. GLM-4-32B-Base-0414 was pre-tra…

Open Source 9.0B ↓ 16.6K
MiMo-7B-Base logo
MiMo-7B-Base
XiaomiMiMo

━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━ Unlocking the Reasoning Potential of Language Model From Pretraining to Posttraining ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━

Open Source 7.0B ↓ 16.2K
Yi-1.5-34B-Chat logo
Yi-1.5-34B-Chat
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 34.0B ↓ 15.6K
Yi-1.5-9B logo
Yi-1.5-9B
01-ai

🐙 GitHub • 👾 Discord • 🐤 Twitter • 💬 WeChat 📝 Paper • 💪 Tech Blog • 🙌 FAQ • 📗 Learning Hub

Open Source 9.0B ↓ 14.8K
Olmo-Hybrid-7B logo
Olmo-Hybrid-7B
allenai

We expand on our Olmo model series by introducing Olmo Hybrid, a new 7B hybrid RNN model in the Olmo family. Olmo Hybrid dramatically outperforms Olmo 3 in final performance, consistently showing roughly 2x data efficiency on core evals over the course of our pretraining run. We…

Open Source 7.0B ↓ 14.6K
LFM2.5-350M-Base logo
LFM2.5-350M-Base
LiquidAI

LFM2.5 is a new family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source 0.35B ↓ 14.4K