AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,109 models for "Chat" Compare
🤖
Tencent-Hunyuan-Large
tencent

&nbsp GITHUB &nbsp&nbsp &nbsp&nbsp🖥️&nbsp&nbsp official website &nbsp&nbsp|&nbsp&nbsp🕖&nbsp&nbsp HunyuanAPI |&nbsp&nbsp🐳&nbsp&nbsp Gitee Technical Report &nbsp&nbsp|&nbsp&nbsp Demo &nbsp&nbsp&nbsp|&nbsp&nbsp Tencent Cloud TI &nbsp&nbsp&nbsp

Open Source ↓ 405
🤖
Yi-VL-34B
01-ai

🤗 Hugging Face • 🤖 ModelScope • 🟣 wisemodel

Multimodal 34.0B ↓ 403
🤖
Nous-Hermes-13b
NousResearch

Nous-Hermes-13b is a state-of-the-art language model fine-tuned on over 300,000 instructions. This model was fine-tuned by Nous Research, with Teknium and Karan4D leading the fine tuning process and dataset curation, Redmond AI sponsoring the compute, and several other contributo…

Open Source 13.0B ↓ 402
🤖
xLAM-7b-fc-r
Salesforce

[Homepage] [Paper] [Discord] [Dataset] [Github]

Open Source 7.0B ↓ 401
🤖
Seed-Coder-8B-Reasoning
ByteDance-Seed

Introduction We are thrilled to introduce Seed-Coder, a powerful, transparent, and parameter-efficient family of open-source code models at the 8B scale, featuring base, instruct, and reasoning variants. Seed-Coder contributes to promote the evolution of open code models through…

Code 8.0B ↓ 400
🤖
CapRL-Qwen3VL-2B
internlm

CapRL 📖 Paper 🏠 Github 🤗 CapRL Collection 🤗 Daily Paper

Multimodal 2.0B ↓ 398
🤖
Stable-DiffCoder-8B-Instruct
ByteDance-Seed

Introduction We are thrilled to introduce Stable-DiffCoder, which is a strong code diffusion large language model. Built directly on the Seed-Coder architecture, data, and training pipeline, it introduces a block diffusion continual pretraining (CPT) stage with a tailored warmup…

Code 8.0B ↓ 398
🤖
Yi-34B-Chat-8bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 397
🤖
GPT-OSS-Swallow-20B-RL-v0.1-MXFP4
tokyotech-llm

GPT-OSS-Swallow v0.1 is a family of large language models available in 20B and 120B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifiable Rewa…

Open Source 20.0B ↓ 396
🤖
Hunyuan-0.5B-Instruct
tencent

🤗  HuggingFace     🤖  ModelScope     🪡  AngelSlim

Open Source 0.5B ↓ 392
🤖
nomos-1
NousResearch

We release Nomos 1 , a specialization of Qwen/Qwen3-30B-A3B-Thinking-2507 for mathematical problem-solving and proof-writing in natural language. Nomos-1 was trained in collaboration with Hillclimb AI.

Open Source ↓ 391
🤖
internlm-20b
internlm

The Shanghai Artificial Intelligence Laboratory, in collaboration with SenseTime Technology, the Chinese University of Hong Kong, and Fudan University, has officially released the 20 billion parameter pretrained model, InternLM-20B. InternLM-20B was pre-trained on over 2.3T Token…

Open Source 20.0B ↓ 386