AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
SmolVLM-256M-Base
HuggingFaceTB

This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.

Multimodal ↓ 1.2K
🤖
MiniCPM4.1-8B-GPTQ
openbmb

GitHub Repo Technical Report Join Us 👋 Contact us in Discord and WeChat

Open Source 8.0B ↓ 1.2K
🤖
MiniMax-M1-80k
MiniMaxAI

We introduce MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. MiniMax-M1 is powered by a hybrid Mixture-of-Experts (MoE) architecture combined with a lightning attention mechanism. The model is developed based on our previous MiniMax-Text-0…

Open Source ↓ 1.2K
🤖
codegen-350M-nl
Salesforce

CodeGen is a family of autoregressive language models for program synthesis from the paper: A Conversational Paradigm for Program Synthesis by Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, Caiming Xiong. The models are originally releas…

Code ↓ 1.2K
🤖
Falcon3-Mamba-7B-Instruct
tiiuae

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B.

Open Source 7.0B ↓ 1.1K
🤖
deepseek-llm-67b-chat
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 67.0B ↓ 1.1K
🤖
Qwen-SEA-LION-v4-32B-IT
aisingapore

Qwen-SEA-LION-v4-32B-IT (Instruct model)

Open Source 32.0B ↓ 1.1K
🤖
japanese-gpt-1b
rinna

This repository provides a 1.3B-parameter Japanese GPT model. The model was trained by rinna Co., Ltd.

Open Source 1.0B ↓ 1.1K
🤖
InternVL3-78B-AWQ
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 78.0B ↓ 1.1K
🤖
InternVL3-38B-AWQ
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.0B ↓ 1.1K
🤖
Gemma-SEA-LION-v3-9B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 9.0B ↓ 1.1K
🤖
Gemma-2-Llama-Swallow-9b-pt-v0.1
tokyotech-llm

Gemma-2-Llama-Swallow series was built by continual pre-training on the gemma-2 models. Gemma 2 Swallow enhanced the Japanese language capabilities of the original Gemma 2 while retaining the English language capabilities. We use approximately 200 billion tokens that were sampled…

Open Source 9.0B ↓ 1.1K