AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,304 models for "Transformer" Compare
🤖
stablecode-completion-alpha-3b-4k
stabilityai

StableCode-Completion-Alpha-3B-4K is a 3 billion parameter decoder-only code completion model pre-trained on diverse set of programming languages that topped the stackoverflow developer survey.

Code 3.0B ↓ 488
🤖
RedPajama-INCITE-7B-Instruct
togethercomputer

RedPajama-INCITE-7B-Instruct was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research r…

Open Source 7.0B ↓ 482
🤖
Medical-Qwen3-Swallow-8B
tokyotech-llm

Medical-Qwen3-Swallow-8B is a medical-domain language model based on tokyotech-llm/Qwen3-Swallow-8B-RL-v0.2. It is designed to support research and development toward safe and trustworthy AI for Japanese clinical settings.

Open Source 8.0B ↓ 480
🤖
EXAONE-Deep-32B
LGAI-EXAONE

We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research. Evaluation results show that 1) EXAONE Deep 2.4B outperforms other models…

Open Source 32.0B ↓ 479
🤖
Falcon-E-3B-Instruct
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 3.0B ↓ 476
🤖
UI2Code_N
zai-org

UI2Code^N: A Visual Language Model for Test-Time Scalable Interactive UI-to-Code Generation

Code ↓ 470
🤖
Swallow-7b-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 469
🤖
OpenELM-3B
apple

Sachin Mehta, Mohammad Hossein Sekhavat, Qingqing Cao, Maxwell Horton, Yanzi Jin, Chenfan Sun, Iman Mirzadeh, Mahyar Najibi, Dmitry Belenko, Peter Zatloukal, Mohammad Rastegari

Open Source 3.0B ↓ 468
🤖
Gemma-SEA-LION-v3-9B
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Open Source 9.0B ↓ 467
🤖
JudgeLM-7B-v1.0
BAAI

--- inference: false language: - en tags: - instruction-finetuning pretty name: JudgeLM-100K task categories: - text-generation ---

Open Source 7.0B ↓ 466
🤖
Baichuan-M3-235B
baichuan-inc

From Inquiry to Decision: Building Trustworthy Medical AI

Open Source 235.0B ↓ 464
🤖
OLMo-Bitnet-1B
NousResearch

OLMo-Bitnet-1B is a 1B parameter model trained using the method described in The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits.

Open Source 1.0B ↓ 456