AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,109 models for "Chat" Compare
🤖
MiMo-V2.5-Pro-Base
XiaomiMiMo

🤗 HuggingFace   📰 Blog   🎨 Xiaomi MiMo API Platform   🗨️ Xiaomi MiMo Studio  

Open Source ↓ 430
🤖
stablelm-2-12b-chat
stabilityai

Stable LM 2 12B Chat is a 12 billion parameter instruction tuned language model trained on a mix of publicly available datasets and synthetic datasets, utilizing Direct Preference Optimization (DPO).

Open Source 12.0B ↓ 428
🤖
Falcon-H1-Tiny-Tool-Calling-90M
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source ↓ 426
🤖
StableBeluga-13B
stabilityai

Use Stable Chat (Research Preview) to test Stability AI's best language models for free

Open Source 13.0B ↓ 424
🤖
Aquila2-7B
BAAI

We opensource our Aquila2 series, now including Aquila2 , the base language models, namely Aquila2-7B and Aquila2-34B , as well as AquilaChat2 , the chat models, namely AquilaChat2-7B and AquilaChat2-34B , as well as the long-text chat models, namely AquilaChat2-7B-16k and Aquila…

Open Source 7.0B ↓ 424
🤖
Falcon-H1-Tiny-90M-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source ↓ 420
🤖
xgen-mm-phi3-mini-instruct-interleave-r-v1.5
Salesforce

Model description xGen-MM is a series of the latest foundational Large Multimodal Models (LMMs) developed by Salesforce AI Research. This series advances upon the successful designs of the BLIP series, incorporating fundamental enhancements that ensure a more robust and superior…

Open Source ↓ 416
🤖
cosmo-1b
HuggingFaceTB

Model Summary This is a 1.8B model trained on Cosmopedia synthetic dataset.

Code 1.0B ↓ 415
🤖
Yi-6B-Chat-4bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 412
🤖
LLaDA-MoE-7B-A1B-Base
inclusionAI

This model is based on the principles described in the paper Large Language Diffusion Models.

Open Source 7.0B ↓ 411
🤖
internlm2_5-7b-chat-1m
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 409
🤖
Llama-3.3-Swallow-70B-Instruct-v0.4
tokyotech-llm

Llama 3.3 Swallow is a large language model (70B) that was built by continual pre-training on the Meta Llama 3.3 model. Llama 3.3 Swallow enhanced the Japanese language capabilities of the original Llama 3.3 while retaining the English language capabilities. We use approximately…

Open Source 70.0B ↓ 408