AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
Mage-VL logo
Mage-VL
microsoft

Mage-VL An Efficient Codec-Native Streaming Multimodal Foundation Model

Multimodal 4.0B ↓ 117.8K
Llama-2-13b-chat-hf logo
Llama-2-13b-chat-hf
meta-llama

Open Source 13.0B ↓ 117.4K
InternVL3-1B-hf logo
InternVL3-1B-hf
OpenGVLab

InternVL3-1B Transformers 🤗 Implementation

Multimodal 0.94B ↓ 112.2K
InternVL3-8B logo
InternVL3-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 7.94B ↓ 110.4K
SmolLM3-3B-Base logo
SmolLM3-3B-Base
HuggingFaceTB

1. Model Summary 2. How to use 3. Evaluation 4. Training 5. Limitations 6. License

Open Source 3.0B ↓ 103.5K
Meta-Llama-3-8B-Instruct logo
Meta-Llama-3-8B-Instruct
NousResearch

Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the av…

Open Source 8.0B ↓ 102.8K
InternVL3-1B logo
InternVL3-1B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 0.94B ↓ 99.9K
DeepSeek-R1-0528-NVFP4-v2 logo
DeepSeek-R1-0528-NVFP4-v2
nvidia

Description: The NVIDIA DeepSeek-R1-0528-FP4 v2 model is the quantized version of the DeepSeek AI's DeepSeek R1 0528 model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA DeepSeek R1…

Reasoning ↓ 87.6K
internlm2-chat-7b logo
internlm2-chat-7b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 86.3K
LFM2.5-230M logo
LFM2.5-230M
LiquidAI

LFM2.5 is a family of hybrid models designed for on-device deployment . It builds on the LFM2 architecture with extended pre-training and reinforcement learning.

Open Source ↓ 81.5K
InternVL3-8B-hf logo
InternVL3-8B-hf
OpenGVLab

InternVL3-8B Transformers 🤗 Implementation

Multimodal 8.0B ↓ 74.6K
Hunyuan-A13B-Instruct logo
Hunyuan-A13B-Instruct
tencent

🤗  Hugging Face       🖥️  Official Website       🕖  HunyuanAPI       🕹️  Demo       🤖  ModelScope

Open Source 13.0B ↓ 72K