AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,294 models for "Transformer" Compare
🤖
Qwen3.8-27B
Qwen

[!Note] This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, TokenSpeed, etc.

Open Source 27.0B ★ 52.0 ↓ 4.7M
🤖
DeepSeek-V4-Flash-0731
deepseek-ai

DeepSeek-V4-Flash-0731 is the official release of DeepSeek-V4-Flash , superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as DeepSeek-V4-Flash-DSpark, i.e. it comes with a speculative decoding module attached.

Open Source ★ 52.0 ↓ 4.6M
🤖
DeepSeek-V4-Flash
deepseek-ai

DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence

Open Source ★ 52.0 ↓ 1.7M
🤖
DeepSeek-V4-Flash-DSpark
deepseek-ai

DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence

Open Source ★ 52.0 ↓ 381.7K
🤖
DeepSeek-V4-Flash-NVFP4
nvidia

Description: The NVIDIA DeepSeek-V4-Flash-NVFP4 model is a quantized version of DeepSeek AI's DeepSeek-V4-Flash model, an autoregressive Mixture-of-Experts language model that uses an optimized Transformer architecture with hybrid attention (Compressed Sparse Attention and Heavil…

Open Source ★ 52.0 ↓ 253.4K
🤖
MiniMax-M3-MXFP8
MiniMaxAI

MiniMax-M3 is a native multimodal model with 1M context. It has ~428B parameters and ~23B activated parameters.

Open Source ★ 45.0 ↓ 363.6K
🤖
MiniMax-M3
MiniMaxAI

MiniMax-M3 is a native multimodal model with 1M context. It has ~428B parameters and ~23B activated parameters.

Open Source ★ 45.0 ↓ 194.5K
🤖
MiniMax-M3-NVFP4
nvidia

Description MiniMax-M3 is a multimodal model with frontier-level coding and agentic capabilities, built on a Mixture-of-Experts architecture with a 1M-token context window. The model processes text, image, video, and computer use inputs and produces text outputs, with emphasis on…

Open Source ★ 45.0 ↓ 186.5K
🤖
Kimi-K2.7-Code-NVFP4
nvidia

Description: The NVIDIA Kimi-K2.7-Code NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2.7-Code model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2.7-Code NV…

Code ★ 43.0 ↓ 489.3K
🤖
Kimi-K2.7-Code
moonshotai

Kimi K2.7 Code is a coding-focused agentic model built upon Kimi K2.6. With substantial improvements on real-world long-horizon coding tasks, it strengthens end-to-end task completion across complex software engineering workflows while improving token efficiency, reducing thinkin…

Code ★ 43.0 ↓ 229.3K
🤖
Hy3-preview
tencent

           

Open Source ★ 42.0 ↓ 61.3K
🤖
Hy3-FP8
tencent

           

Open Source ★ 42.0 ↓ 39K