AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
StableBeluga-13B
stabilityai

Use Stable Chat (Research Preview) to test Stability AI's best language models for free

Open Source 13.0B ↓ 424
🤖
Aquila2-7B
BAAI

We opensource our Aquila2 series, now including Aquila2 , the base language models, namely Aquila2-7B and Aquila2-34B , as well as AquilaChat2 , the chat models, namely AquilaChat2-7B and AquilaChat2-34B , as well as the long-text chat models, namely AquilaChat2-7B-16k and Aquila…

Open Source 7.0B ↓ 424
🤖
aya-vision-32b
CohereLabs

Multimodal 32.0B ↓ 423
🤖
Falcon-H1-Tiny-90M-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source ↓ 420
🤖
xgen-mm-phi3-mini-instruct-interleave-r-v1.5
Salesforce

Model description xGen-MM is a series of the latest foundational Large Multimodal Models (LMMs) developed by Salesforce AI Research. This series advances upon the successful designs of the BLIP series, incorporating fundamental enhancements that ensure a more robust and superior…

Open Source ↓ 416
🤖
cosmo-1b
HuggingFaceTB

Model Summary This is a 1.8B model trained on Cosmopedia synthetic dataset.

Code 1.0B ↓ 415
🤖
Yi-6B-Chat-4bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 412
🤖
open-calm-medium
cyberagent

OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.

Open Source ↓ 411
🤖
LLaDA-MoE-7B-A1B-Base
inclusionAI

This model is based on the principles described in the paper Large Language Diffusion Models.

Open Source 7.0B ↓ 411
🤖
MiMo-Embodied-7B
XiaomiMiMo

🤗 HuggingFace   📔 Technical Report  

Open Source 7.0B ↓ 409
🤖
internlm2_5-7b-chat-1m
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 409
🤖
Llama-3.3-Swallow-70B-Instruct-v0.4
tokyotech-llm

Llama 3.3 Swallow is a large language model (70B) that was built by continual pre-training on the Meta Llama 3.3 model. Llama 3.3 Swallow enhanced the Japanese language capabilities of the original Llama 3.3 while retaining the English language capabilities. We use approximately…

Open Source 70.0B ↓ 408