AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,309 models for "Transformer" Compare
🤖
AI21-Jamba2-Mini
ai21labs

Jamba2 Mini is an open source small language model built for enterprise reliability. With 12B active parameters (52B total), it delivers precise question answering without the computational overhead of reasoning models. The model's SSM-Transformer architecture provides a memory-e…

Open Source ↓ 1.3K
🤖
Falcon3-7B-Instruct-1.58bit
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 7.0B ↓ 1.3K
🤖
stablelm-tuned-alpha-3b
stabilityai

StableLM-Tuned-Alpha is a suite of 3B and 7B parameter decoder-only language models built on top of the StableLM-Base-Alpha models and further fine-tuned on various chat and instruction-following datasets.

Open Source 3.0B ↓ 1.3K
🤖
GLM-4.5-FP8
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.5 technical blog . 📍 Use GLM-4.5 API services on Z.ai API Platform (Global) or Zhipu AI Open Platform (Mainland China) . 👉 One click to GLM-4.5 .

Open Source ↓ 1.3K
🤖
xLAM-2-1b-fc-r
Salesforce

Large Action Models (LAMs) are advanced language models designed to enhance decision-making by translating user intentions into executable actions. As the brains of AI agents , LAMs autonomously plan and execute tasks to achieve specific goals, making them invaluable for automati…

Open Source 1.0B ↓ 1.3K
🤖
SingGuard-2b
inclusionAI

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning

Open Source 2.0B ↓ 1.3K
🤖
InternVL3_5-30B-A3B-HF
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 30.0B ↓ 1.3K
🤖
deepseek-math-7b-rl
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 7.0B ↓ 1.3K
🤖
bilingual-gpt-neox-4b
rinna

Overview This repository provides an English-Japanese bilingual GPT-NeoX model of 3.8 billion parameters.

Open Source 4.0B ↓ 1.3K
🤖
Falcon-H1-3B-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 3.0B ↓ 1.2K
🤖
Falcon-H1-34B-Instruct
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 34.0B ↓ 1.2K
🤖
DeepHermes-3-Llama-3-8B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 8.0B ↓ 1.2K