AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
internlm2_5-1_8b-chat logo
internlm2_5-1_8b-chat
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 1.8B ↓ 3.2K
Apriel-5B-Instruct logo
Apriel-5B-Instruct
ServiceNow-AI

1. Model Summary 2. Evaluation 3. Intended Use 4. Limitations 5. Security and Responsible Use 6. License 7. Citation

Open Source 5.0B ↓ 3.1K
stablelm-base-alpha-3b logo
stablelm-base-alpha-3b
stabilityai

📢 DISCLAIMER : The StableLM-Base-Alpha models have been superseded. Find the latest versions in the Stable LM Collection here.

Open Source 3.64B ↓ 3K
internlm-7b logo
internlm-7b
internlm

InternLM has open-sourced a 7 billion parameter base model tailored for practical scenarios. The model has the following characteristics: - It leverages trillions of high-quality tokens for training to establish a powerful knowledge base. - It provides a versatile toolset for use…

Open Source 7.0B ↓ 2.8K
InternVL3-38B-hf logo
InternVL3-38B-hf
OpenGVLab

InternVL3-38B Transformers 🤗 Implementation

Multimodal 38.4B ↓ 2.8K
sarvam-m logo
sarvam-m
sarvamai

sarvam-m is a multilingual, hybrid-reasoning, text-only language model built on Mistral-Small. This post-trained version delivers exceptional improvements over the base model:

Open Source 24.0B ↓ 2.7K
evo-1-8k-base logo
evo-1-8k-base
togethercomputer

We identified and fixed an issue related to a wrong permutation of some projections, which affects generation quality. To use the new model revision, please load as follows:

Open Source 7.0B ↓ 2.7K
Ling-flash-2.0 logo
Ling-flash-2.0
inclusionAI

🤗 Hugging Face &nbsp&nbsp &nbsp&nbsp🤖 ModelScope

Open Source 100.0B ↓ 2.5K
LFM2-VL-3B logo
LFM2-VL-3B
LiquidAI

LFM2-VL-3B is the newest and most capable model in Liquid AI's multimodal LFM2-VL series, designed to process text and images with variable resolutions. Built on the LFM2 backbone, it extends the architecture for higher-capacity reasoning and stronger visual understanding while r…

Multimodal 3.0B ↓ 2.3K
glm-4-9b-hf logo
glm-4-9b-hf
zai-org

If you are using the weights from this repository, please update to

Open Source 9.0B ↓ 2.3K
Llama-3.3-Swallow-70B-Instruct-v0.4 logo
Llama-3.3-Swallow-70B-Instruct-v0.4
tokyotech-llm

Llama 3.3 Swallow is a large language model (70B) that was built by continual pre-training on the Meta Llama 3.3 model. Llama 3.3 Swallow enhanced the Japanese language capabilities of the original Llama 3.3 while retaining the English language capabilities. We use approximately…

Open Source 70.0B ↓ 2.2K
Hunyuan-A13B-Pretrain logo
Hunyuan-A13B-Pretrain
tencent

🤗  Hugging Face       🖥️  Official Website       🕖  HunyuanAPI       🕹️  Demo       🤖  ModelScope

Open Source 13.0B ↓ 2.1K