AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

562 models for "Vision" Compare
SmolLM-135M-Instruct logo
SmolLM-135M-Instruct
HuggingFaceTB

SmolLM is a series of small language models available in three sizes: 135M, 360M, and 1.7B parameters.

Open Source 0.135B ↓ 26.9K
InternVL3_5-8B-HF logo
InternVL3_5-8B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 8.5B ↓ 26.5K
SmolVLM-Instruct logo
SmolVLM-Instruct
HuggingFaceTB

SmolVLM is a compact open multimodal model that accepts arbitrary sequences of image and text inputs to produce text outputs. Designed for efficiency, SmolVLM can answer questions about images, describe visual content, create stories grounded on multiple images, or function as a…

Multimodal 2.0B ↓ 25.8K
LFM2.5-VL-3B logo
LFM2.5-VL-3B
LiquidAI

LFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment . It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both text and images, and uses the LFM2.5-2.6B language model as its backbone, combi…

Multimodal 3.0B ↓ 25K
Mini-InternVL-Chat-2B-V1-5 logo
Mini-InternVL-Chat-2B-V1-5
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 2.2B ↓ 24K
MiniCPM-Llama3-V-2_5 logo
MiniCPM-Llama3-V-2_5
openbmb

A GPT-4V Level Multimodal LLM on Your Phone

Multimodal 8.5B ↓ 22.7K
InternVL2-8B logo
InternVL2-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 8.1B ↓ 22K
granite-3.0-8b-instruct logo
granite-3.0-8b-instruct
ibm-granite

Model Summary: Granite-3.0-8B-Instruct is a 8B parameter model finetuned from Granite-3.0-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. This model is developed using a diverse set of techniques…

Open Source 8.1B ↓ 21.8K
Intern-S1 logo
Intern-S1
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat

Open Source ↓ 20.3K
OLMoE-1B-7B-0924-Instruct logo
OLMoE-1B-7B-0924-Instruct
allenai

OLMoE-1B-7B-Instruct is a Mixture-of-Experts LLM with 1B active and 7B total parameters released in September 2024 (0924) that has been adapted via SFT and DPO from OLMoE-1B-7B. It yields state-of-the-art performance among models with a similar cost (1B) and is competitive with m…

Open Source 1.0B ↓ 19.4K
InternVL3_5-4B logo
InternVL3_5-4B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.7B ↓ 19.4K
ERNIE-4.5-21B-A3B-PT logo
ERNIE-4.5-21B-A3B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Open Source 21.0B ↓ 19.1K