AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,314 models for "Transformer" Compare
🤖
blip2-opt-6.7b
Salesforce

BLIP-2 model, leveraging OPT-6.7b (a large language model with 6.7 billion parameters). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Open Source 6.7B ↓ 23.9K
🤖
Baichuan-7B
baichuan-inc

Baichuan-7B是由百川智能开发的一个开源的大规模预训练模型。基于Transformer结构,在大约1.2万亿tokens上训练的70亿参数模型,支持中英双语,上下文窗口长度为4096。在标准的中文和英文权威benchmark(C-EVAL/MMLU)上均取得同尺寸最好的效果。

Open Source 7.0B ↓ 23.5K
🤖
Qwen-SEA-LION-v4-8B-VL
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Multimodal 8.0B ↓ 23.4K
🤖
codegen2-16B_P
Salesforce

CodeGen2 is a family of autoregressive language models for program synthesis , introduced in the paper:

Code 16.0B ↓ 23.3K
🤖
internlm2-7b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 23.1K
🤖
SOLAR-0-70b-16bit
upstage

Updates Solar, a new bot created by Upstage, is now available on Poe . As a top-ranked model on the HuggingFace Open LLM leaderboard, and a fine tune of Llama 2, Solar is a great example of the progress enabled by open source. Try now at https://poe.com/Solar-0-70b

Open Source 70.0B ↓ 23.1K
🤖
LFM2.5-VL-3B
LiquidAI

LFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment . It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both text and images, and uses the LFM2.5-2.6B language model as its backbone, combi…

Multimodal 3.0B ↓ 22.8K
🤖
GLM-Z1-32B-0414
zai-org

The GLM family welcomes a new generation of open-source models, the GLM-4-32B-0414 series, featuring 32 billion parameters. Its performance is comparable to OpenAI's GPT series and DeepSeek's V3/R1 series, and it supports very user-friendly local deployment features. GLM-4-32B-Ba…

Reasoning 32.0B ↓ 22.7K
🤖
Nous-Hermes-llama-2-7b
NousResearch

Compute provided by our project sponsor Redmond AI, thank you! Follow RedmondAI on Twitter @RedmondAI.

Open Source 7.0B ↓ 21.8K
🤖
Olmo-Hybrid-7B
allenai

We expand on our Olmo model series by introducing Olmo Hybrid, a new 7B hybrid RNN model in the Olmo family. Olmo Hybrid dramatically outperforms Olmo 3 in final performance, consistently showing roughly 2x data efficiency on core evals over the course of our pretraining run. We…

Open Source 7.0B ↓ 21.3K
🤖
GLM-4.7-FP8
zai-org

👋 Join our Discord community. 📖 Check out the GLM-4.7 technical blog , technical report(GLM-4.5) . 📍 Use GLM-4.7 API services on Z.ai API Platform. 👉 One click to GLM-4.7 .

Open Source ↓ 20.7K
🤖
InternVL2_5-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 Mini-InternVL\]](https://arxiv.org/abs/2410.16261) [\[📜 InternVL 2.5\]](https://huggingface.co/…

Multimodal 8.08B ↓ 20.4K