AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,108 models for "Chat" Compare
🤖
Qwen3.6-35B-A3B-FP8
Qwen

[!Note] This repository contains FP8-quantized model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc. The quantization method is fin…

Open Source 35.0B ★ 32.0 ↓ 13.3M
🤖
Qwen3.6-35B-A3B-NVFP4
nvidia

Description: The NVIDIA Qwen3.6-35B-A3B-NVFP4 model is the quantized version of Alibaba's Qwen3.6-35B-A3B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.6-35B-A3B-NVFP4 m…

Open Source 35.0B ★ 32.0 ↓ 10.8M
🤖
Qwen3.6-35B-A3B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Open Source 35.0B ★ 32.0 ↓ 4.9M
🤖
Ring-2.6-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 ling.tbox.cn       Tech Report

Open Source ★ 32.0 ↓ 567
🤖
Step-3.7-Flash
stepfun-ai

[ModelPage] : https://static.stepfun.com/blog/step-3.7-flash/

Open Source ★ 31.0 ↓ 38.8K
🤖
Olmo-3-32B-Think
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 31.0 ↓ 9.7K
🤖
Step-3.7-Flash-FP8
stepfun-ai

[ModelPage] : https://static.stepfun.com/blog/step-3.7-flash/

Open Source ★ 31.0 ↓ 8.7K
🤖
Step-3.7-Flash-NVFP4
stepfun-ai

[ModelPage] : https://static.stepfun.com/blog/step-3.7-flash/

Open Source ★ 31.0 ↓ 5.4K
🤖
K-EXAONE-2.0-750B-A37B
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 4.1K
🤖
Olmo-3-32B-Think-DPO
allenai

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Open Source 32.0B ★ 31.0 ↓ 3.9K
🤖
K-EXAONE-2.0-750B-A37B-NVFP4
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 1.9K
🤖
K-EXAONE-2.0-750B-A37B-FP8
LGAI-EXAONE

We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…

Open Source 750.0B ★ 31.0 ↓ 1.9K