AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,114 models for "Chat" Compare
🤖
InternVL3-1B-hf
OpenGVLab

InternVL3-1B Transformers 🤗 Implementation

Multimodal 0.94B ↓ 143.3K
🤖
NVIDIA-Nemotron-Parse-v1.1
nvidia

NVIDIA Nemotron Parse v1.1 is designed to understand document semantics and extract text and tables elements with spatial grounding. Given an image, NVIDIA Nemotron Parse v1.1 produces structured annotations, including formatted text, bounding-boxes and the corresponding semantic…

Multimodal 0.885B ↓ 141K
🤖
EXAONE-3.5-32B-Instruct-AWQ
LGAI-EXAONE

We introduce EXAONE 3.5, a collection of instruction-tuned bilingual (English and Korean) generative models ranging from 2.4B to 32B parameters, developed and released by LG AI Research. EXAONE 3.5 language models include: 1) 2.4B model optimized for deployment on small or resour…

Open Source 32.0B ↓ 135.7K
🤖
Phi-3.5-MoE-instruct
microsoft

Phi-3.5-MoE is a lightweight, state-of-the-art open model built upon datasets used for Phi-3 - synthetic data and filtered publicly available documents - with a focus on very high-quality, reasoning dense data. The model supports multilingual and comes with 128K context length (i…

Open Source ↓ 134.1K
🤖
blip2-flan-t5-xl
Salesforce

BLIP-2 model, leveraging Flan T5-xl (a large language model). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Multimodal 4.0B ↓ 133K
🤖
Qwen3-8B-NVFP4
nvidia

Description: The NVIDIA Qwen3-8B FP4 model is the quantized version of Alibaba's Qwen3-8B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3-8B FP4 model is quantized with Te…

Open Source 8.0B ↓ 131.4K
🤖
paligemma-3b-mix-224
google

Multimodal 2.92B ↓ 117K
🤖
GLM-4.5-Air
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.5 technical blog , technical report , and Zhipu AI technical documentation . 📍 Use GLM-4.5 API services on Z.ai API Platform (Global) or Zhipu AI Open Platform (Mainland China) . 👉 One click to GLM-4.5 .

Open Source ↓ 116.4K
🤖
Llama-2-13b-chat-hf
meta-llama

Open Source 13.0B ↓ 116.3K
🤖
gemma-1.1-2b-it
google

Open Source 2.0B ↓ 115.7K
🤖
LocateAnything-3B
nvidia

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

Multimodal 3.0B ↓ 115.1K
🤖
granite-vision-4.1-4b
ibm-granite

Model Summary: Granite Vision 4.1 4B is a vision-language model (VLM) that delivers frontier-level performance on structured document extraction tasks — chart extraction, table extraction, and semantic key-value pair extraction — in a compact 4B parameter footprint, providing a l…

Multimodal 4.0B ↓ 110.9K