AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
Tencent-Hunyuan-Large
tencent

&nbsp GITHUB &nbsp&nbsp &nbsp&nbsp🖥️&nbsp&nbsp official website &nbsp&nbsp|&nbsp&nbsp🕖&nbsp&nbsp HunyuanAPI |&nbsp&nbsp🐳&nbsp&nbsp Gitee Technical Report &nbsp&nbsp|&nbsp&nbsp Demo &nbsp&nbsp&nbsp|&nbsp&nbsp Tencent Cloud TI &nbsp&nbsp&nbsp

Open Source ↓ 405
🤖
Yi-VL-34B
01-ai

🤗 Hugging Face • 🤖 ModelScope • 🟣 wisemodel

Multimodal 34.0B ↓ 403
🤖
Nous-Hermes-13b
NousResearch

Nous-Hermes-13b is a state-of-the-art language model fine-tuned on over 300,000 instructions. This model was fine-tuned by Nous Research, with Teknium and Karan4D leading the fine tuning process and dataset curation, Redmond AI sponsoring the compute, and several other contributo…

Open Source 13.0B ↓ 402
🤖
xLAM-7b-fc-r
Salesforce

[Homepage] [Paper] [Discord] [Dataset] [Github]

Open Source 7.0B ↓ 401
🤖
Seed-Coder-8B-Reasoning
ByteDance-Seed

Introduction We are thrilled to introduce Seed-Coder, a powerful, transparent, and parameter-efficient family of open-source code models at the 8B scale, featuring base, instruct, and reasoning variants. Seed-Coder contributes to promote the evolution of open code models through…

Code 8.0B ↓ 400
🤖
CapRL-Qwen3VL-2B
internlm

CapRL 📖 Paper 🏠 Github 🤗 CapRL Collection 🤗 Daily Paper

Multimodal 2.0B ↓ 398
🤖
TinySolar-248m-4k
upstage

Open Source ↓ 398
🤖
Stable-DiffCoder-8B-Instruct
ByteDance-Seed

Introduction We are thrilled to introduce Stable-DiffCoder, which is a strong code diffusion large language model. Built directly on the Seed-Coder architecture, data, and training pipeline, it introduces a block diffusion continual pretraining (CPT) stage with a tailored warmup…

Code 8.0B ↓ 398
🤖
Llama-SEA-LION-v3-70B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 70.0B ↓ 397
🤖
Yi-34B-Chat-8bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 397
🤖
GPT-OSS-Swallow-20B-RL-v0.1-MXFP4
tokyotech-llm

GPT-OSS-Swallow v0.1 is a family of large language models available in 20B and 120B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifiable Rewa…

Open Source 20.0B ↓ 396
🤖
Hunyuan-0.5B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 0.5B ↓ 392