AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
ContextPilot-E4B logo
ContextPilot-E4B
tencent

ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL

Open Source 8.0B ↓ 764
GLM-4.1V-9B-Base logo
GLM-4.1V-9B-Base
zai-org

📖 View the GLM-4.1V-9B-Thinking paper . 📍 Using GLM-4.1V-9B-Thinking API at Zhipu Foundation Model Open Platform

Open Source 9.0B ↓ 763
UI-Mate-27B logo
UI-Mate-27B
tencent

UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations

Open Source 27.0B ↓ 761
ContextPilot-8B logo
ContextPilot-8B
tencent

ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL

Open Source 8.0B ↓ 749
Gemma-2-Llama-Swallow-9b-it-v0.1 logo
Gemma-2-Llama-Swallow-9b-it-v0.1
tokyotech-llm

Gemma-2-Llama-Swallow series was built by continual pre-training on the gemma-2 models. Gemma 2 Swallow enhanced the Japanese language capabilities of the original Gemma 2 while retaining the English language capabilities. We use approximately 200 billion tokens that were sampled…

Open Source 9.0B ↓ 742
RedPajama-INCITE-7B-Instruct logo
RedPajama-INCITE-7B-Instruct
togethercomputer

RedPajama-INCITE-7B-Instruct was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research r…

Open Source 7.0B ↓ 729
SingGuard-NSFA-4B logo
SingGuard-NSFA-4B
inclusionAI

SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification

Open Source 4.0B ↓ 726
evo-1-131k-base logo
evo-1-131k-base
togethercomputer

We identified and fixed an issue related to a wrong permutation of some projections, which affects generation quality. To use the new model revision, please load as follows:

Open Source 7.0B ↓ 717
Qwen-SEA-LION-v4-8B-VL logo
Qwen-SEA-LION-v4-8B-VL
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Multimodal 8.0B ↓ 715
SynLogic-7B logo
SynLogic-7B
MiniMaxAI

🐙 GitHub Repo: https://github.com/MiniMax-AI/SynLogic 📜 Paper (arXiv): https://arxiv.org/abs/2505.19641 🤗 Dataset: SynLogic on Hugging Face

Open Source 7.0B ↓ 712
InternVL3-78B-AWQ logo
InternVL3-78B-AWQ
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 78.0B ↓ 702
SmolVLM2-2.2B-Base logo
SmolVLM2-2.2B-Base
HuggingFaceTB

This is the base model for SmolVLM2-2.2B, a lightweight multimodal model designed to analyze video content. The model processes videos, images, and text inputs to generate text outputs - whether answering questions about media files, comparing visual content, or transcribing text…

Multimodal 2.2B ↓ 689