AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,296 models for "Transformer" Compare
🤖
JanusCoder-8B
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report

Code 8.0B ↓ 183
🤖
santacoderpack
bigcode

1. Model Summary 2. Use 3. Training 4. Citation

Code ↓ 183
🤖
Redmond-Hermes-Coder
NousResearch

Redmond-Hermes-Coder 15B is a state-of-the-art language model fine-tuned on over 300,000 instructions. This model was fine-tuned by Nous Research, with Teknium and Karan4D leading the fine tuning process and dataset curation, Redmond AI sponsoring the compute, and several other c…

Code ↓ 183
🤖
Qwen-SEA-LION-v4-32B-IT-8BIT
aisingapore

Qwen-SEA-LION-v4-32B-IT-8BIT (GPTQ model)

Open Source 32.0B ↓ 181
🤖
Llama-SEA-Guard-8B-2602
aisingapore

SEA-Safeguard is a collection of safety-focused Large Language Models (LLMs) built upon the SEA-LION family, designed specifically for the Southeast Asia (SEA) region.

Open Source 8.0B ↓ 181
🤖
StripedHyena-Nous-7B
togethercomputer

One of the focus areas at Together Research is new architectures for long context, improved training, and inference performance over the Transformer architecture. Spinning out of a research program from our team and academic collaborators, with roots in signal processing-inspired…

Open Source 7.0B ↓ 180
🤖
JanusCoder-14B
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report

Code 14.0B ↓ 180
🤖
Intern-S1-mini-FP8
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat

Open Source ↓ 180
🤖
japanese-stablelm-instruct-beta-7b
stabilityai

A cute robot wearing a kimono writes calligraphy with one single brush — Stable Diffusion XL

Open Source 7.0B ↓ 178
🤖
Qwen-SEA-LION-v4-32B-IT-4BIT
aisingapore

Qwen-SEA-LION-v4-32B-IT-4BIT (GPTQ model)

Open Source 32.0B ↓ 177
🤖
japanese-stablelm-instruct-alpha-7b-v2
stabilityai

"A parrot able to speak Japanese, ukiyoe, edo period" — Stable Diffusion XL

Open Source 7.0B ↓ 176
🤖
AquilaDense-16B
BAAI

AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]

Open Source 16.0B ↓ 176