AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
Llama-2-13b
meta-llama

Open Source 13.0B ↓ 19
🤖
llama-3-youko-70b-instruct-gptq
rinna

Llama 3 Youko 70B Instruct GPTQ (rinna/llama-3-youko-70b-instruct-gptq)

Open Source 70.0B ↓ 18
🤖
llama-3-youko-8b-gptq
rinna

Llama 3 Youko 8B GPTQ (rinna/llama-3-youko-8b-gptq)

Open Source 8.0B ↓ 17
🤖
qwen2.5-math-rlep
Kwai-Klear

RLEP: Reinforcement Learning with Experience Replay for LLM Reasoning

Open Source ↓ 16
🤖
opencole-typographylmm-llava-v1.5-7b-lora
cyberagent

This model is based on LLaVA1.5-7b. The model is finetuned with LoRA on OpenCOLE1.0 dataset to generate text layouts.

Open Source 7.0B ↓ 16
🤖
Qwen2.5-0.5B-DuDi
aisingapore

This model is a fine-tuned version of Qwen2.5-0.5B. It has been trained using TRL.

Open Source 0.5B ↓ 15
🤖
youri-7b-instruction-gptq
rinna

Overview rinna/youri-7b-instruction-gptq is the quantized model for rinna/youri-7b-instruction using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.

Open Source 7.0B ↓ 12
🤖
youri-7b-gptq
rinna

Overview rinna/youri-7b-gptq is the quantized model for rinna/youri-7b using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.

Open Source 7.0B ↓ 12
🤖
youri-7b-chat-gptq
rinna

Overview rinna/youri-7b-chat-gptq is the quantized model for rinna/youri-7b-chat using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.

Open Source 7.0B ↓ 12
🤖
CodeLlama-70b-Python-hf
meta-llama

Code 70.0B ↓ 12
🤖
Llama-Guard-3-1B-INT4
meta-llama

Open Source 1.0B ↓ 10
🤖
Qwen3-0.6B-Base-DuDi
aisingapore

This model is a fine-tuned version of Qwen3-0.6B-Base. It has been trained using TRL.

Open Source 0.6B ↓ 10