AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,112 models for "Chat" Compare
🤖
SmolLM2-1.7B-Instruct
HuggingFaceTB

1. Model Summary 2. Evaluation 3. Examples 4. Limitations 5. Training 6. License 7. Citation

Open Source 1.7B ↓ 185.2K
🤖
Mistral-7B-Instruct-v0.1
mistralai

py from mistral common.tokens.tokenizers.mistral import MistralTokenizer from mistral common.protocol.instruct.messages import UserMessage from mistral common.protocol.instruct.request import ChatCompletionRequest

Open Source 7.0B ↓ 178.1K
🤖
InternVL2-26B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 25.5B ↓ 177.1K
🤖
Llama-3_3-Nemotron-Super-49B-v1
nvidia

Llama-3.3-Nemotron-Super-49B-v1 is a large language model (LLM) which is a derivative of Meta Llama-3.3-70B-Instruct (AKA the reference model ). It is a reasoning model that is post trained for reasoning, human chat preferences, and tasks, such as RAG and tool calling. The model…

Open Source 49.0B ↓ 176.5K
🤖
DialoGPT-medium
microsoft

A State-of-the-Art Large-scale Pretrained Response generation model (DialoGPT)

Open Source 345.0B ↓ 176.4K
🤖
Kimi-K2-Instruct
moonshotai

📰   Tech Blog         📄   Paper

Open Source ↓ 166.4K
🤖
SmolVLM2-2.2B-Instruct
HuggingFaceTB

SmolVLM2-2.2B is a lightweight multimodal model designed to analyze video content. The model processes videos, images, and text inputs to generate text outputs - whether answering questions about media files, comparing visual content, or transcribing text from images. Despite its…

Multimodal 2.2B ↓ 162.7K
🤖
Nemotron-Labs-Diffusion-8B
nvidia

Nemotron-Labs-Diffusion is a tri-mode language model that supports both AR decoding and diffusion-based parallel decoding by simply switching the attention pattern of the same model during inference. The synergy between these two modes enables a third mode, called self-speculatio…

Open Source 8.0B ↓ 160.2K
🤖
Step-3.5-Flash
stepfun-ai

Step 3.5 Flash (visit website) is our most capable open-source foundation model, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B p…

Open Source ↓ 155.3K
🤖
granite-3.0-8b-instruct
ibm-granite

Model Summary: Granite-3.0-8B-Instruct is a 8B parameter model finetuned from Granite-3.0-8B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. This model is developed using a diverse set of techniques…

Open Source 8.1B ↓ 152.6K
🤖
InternVL3-1B-hf
OpenGVLab

InternVL3-1B Transformers 🤗 Implementation

Multimodal 0.94B ↓ 143.3K
🤖
NVIDIA-Nemotron-Parse-v1.1
nvidia

NVIDIA Nemotron Parse v1.1 is designed to understand document semantics and extract text and tables elements with spatial grounding. Given an image, NVIDIA Nemotron Parse v1.1 produces structured annotations, including formatted text, bounding-boxes and the corresponding semantic…

Multimodal 0.885B ↓ 141K