AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,112 models for "Chat" Compare
🤖
Gemma-SEA-LION-v4-27B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 27.0B ↓ 859
🤖
Bunny-v1_0-3B
BAAI

This is the merged weights of bunny-phi-2-siglip-lora.

Open Source 3.0B ↓ 854
🤖
deepseek-coder-7b-base-v1.5
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 7.0B ↓ 850
🤖
blip2-flan-t5-xxl
Salesforce

BLIP-2 model, leveraging Flan T5-xxl (a large language model). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Open Source ↓ 850
🤖
LLaDA2.0-mini-CAP
inclusionAI

LLaDA2.0-mini-CAP is an enhanced version of LLaDA2.0-mini that incorporates Confidence-Aware Parallel (CAP) Training for significantly improved inference efficiency. Built upon the 16B-A1B Mixture-of-Experts (MoE) diffusion architecture, this model achieves faster parallel decodi…

Open Source ↓ 832
🤖
EXAONE-Deep-2.4B
LGAI-EXAONE

We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research. Evaluation results show that 1) EXAONE Deep 2.4B outperforms other models…

Open Source 2.4B ↓ 817
🤖
RedPajama-INCITE-7B-Base
togethercomputer

RedPajama-INCITE-7B-Base was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research resea…

Open Source 7.0B ↓ 804
🤖
MathForm-8B
openbmb

MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verification-Guided Refinement

Open Source 8.0B ↓ 797
🤖
CoDA-v0-Base
Salesforce

Try CoDA · Paper · Model Collection · GitHub Repository

Open Source ↓ 788
🤖
Emu3-Gen-hf
BAAI

Emu3: Next-Token Prediction is All You Need

Open Source ↓ 787
🤖
llama-30b-instruct-2048
upstage

Developed by : Upstage Backbone Model : LLaMA Variations : It has different model parameter sizes and sequence lengths: 30B/1024, 30B/2048, 65B/1024 Language(s) : English Library : HuggingFace Transformers License : This model is under a Non-commercial Bespoke License and governe…

Open Source 30.0B ↓ 781
🤖
Emu3-Chat
BAAI

Emu3: Next-Token Prediction is All You Need

Open Source ↓ 770