AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
Qwen2.5-0.5B logo
Qwen2.5-0.5B
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 0.5B ↓ 1.5M
Llama-3.2-3B-Instruct logo
Llama-3.2-3B-Instruct
meta-llama

Open Source 3.21B ↓ 1.4M
Qwen2-0.5B logo
Qwen2-0.5B
Qwen

Qwen2 is the new series of Qwen large language models. For Qwen2, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters, including a Mixture-of-Experts model. This repo contains the 0.5B Qwen2 base language mod…

Open Source 0.5B ↓ 830.7K
Phi-3.5-vision-instruct logo
Phi-3.5-vision-instruct
microsoft

Phi-3.5-vision is a lightweight, state-of-the-art open multimodal model built upon datasets which include - synthetic data and filtered publicly available websites - with a focus on very high-quality, reasoning dense data both on text and vision. The model belongs to the Phi-3 mo…

Multimodal 4.2B ↓ 740.1K
Qwen2.5-1.5B logo
Qwen2.5-1.5B
Qwen

Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2:

Open Source 1.5B ↓ 653K
Llama-3.1-70B-Instruct logo
Llama-3.1-70B-Instruct
meta-llama

Open Source 70.0B ↓ 391.2K
Llama-2-7b-chat-hf logo
Llama-2-7b-chat-hf
meta-llama

Open Source 7.0B ↓ 347.5K
Meta-Llama-3.1-8B-Instruct logo
Meta-Llama-3.1-8B-Instruct
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 8.0B ↓ 225.5K
granite-guardian-3.3-8b logo
granite-guardian-3.3-8b
ibm-granite

Model Summary: Granite Guardian 3.3 8b is a specialized Granite 3.3 8B model designed to judge if the input prompts and the output responses of an LLM based system meet specified criteria. The model comes pre-baked with certain criteria including but not limited to: jailbreak att…

Open Source 8.0B ↓ 214K
Meta-Llama-3.1-8B logo
Meta-Llama-3.1-8B
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 8.0B ↓ 203.1K
Llama-2-7b-hf logo
Llama-2-7b-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 7B pretrained model, converted for the Hugging Face Transformers format. Links to other models can be found…

Open Source 7.0B ↓ 138.6K
Llama-2-13b-chat-hf logo
Llama-2-13b-chat-hf
meta-llama

Open Source 13.0B ↓ 117.4K