AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,114 models for "Chat" Compare
🤖
AI21-Jamba2-Mini
ai21labs

Jamba2 Mini is an open source small language model built for enterprise reliability. With 12B active parameters (52B total), it delivers precise question answering without the computational overhead of reasoning models. The model's SSM-Transformer architecture provides a memory-e…

Open Source ↓ 1.3K
🤖
Falcon3-7B-Instruct-1.58bit
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 7.0B ↓ 1.3K
🤖
stablelm-tuned-alpha-3b
stabilityai

StableLM-Tuned-Alpha is a suite of 3B and 7B parameter decoder-only language models built on top of the StableLM-Base-Alpha models and further fine-tuned on various chat and instruction-following datasets.

Open Source 3.0B ↓ 1.3K
🤖
GLM-4.5-FP8
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.5 technical blog . 📍 Use GLM-4.5 API services on Z.ai API Platform (Global) or Zhipu AI Open Platform (Mainland China) . 👉 One click to GLM-4.5 .

Open Source ↓ 1.3K
🤖
xLAM-2-1b-fc-r
Salesforce

Large Action Models (LAMs) are advanced language models designed to enhance decision-making by translating user intentions into executable actions. As the brains of AI agents , LAMs autonomously plan and execute tasks to achieve specific goals, making them invaluable for automati…

Open Source 1.0B ↓ 1.3K
🤖
SingGuard-2b
inclusionAI

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

Open Source 2.0B ↓ 1.3K
🤖
InternVL3_5-30B-A3B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 30.0B ↓ 1.3K
🤖
deepseek-math-7b-rl
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 7.0B ↓ 1.3K
🤖
GPT-OSS-Swallow-120B-RL-v0.1
tokyotech-llm

GPT-OSS-Swallow v0.1 is a family of large language models available in 20B and 120B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifiable Rewa…

Open Source 120.0B ↓ 1.2K
🤖
DeepHermes-3-Llama-3-8B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 8.0B ↓ 1.2K
🤖
GLM-Z1-9B-0414
zai-org

The GLM family welcomes a new generation of open-source models, the GLM-4-32B-0414 series, featuring 32 billion parameters. Its performance is comparable to OpenAI's GPT series and DeepSeek's V3/R1 series, and it supports very user-friendly local deployment features. GLM-4-32B-Ba…

Open Source 9.0B ↓ 1.2K
🤖
MiniCPM4.1-8B-GPTQ
openbmb

GitHub Repo Technical Report Join Us 👋 Contact us in Discord and WeChat

Open Source 8.0B ↓ 1.2K