AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
Meta-Llama-3-8B-Instruct logo
Meta-Llama-3-8B-Instruct
NousResearch

Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the av…

Open Source 8.0B ↓ 102.8K
gemma-1.1-2b-it logo
gemma-1.1-2b-it
google

Open Source 2.0B ↓ 100.3K
MiniCPM-V-4_5 logo
MiniCPM-V-4_5
openbmb

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

Multimodal 8.7B ↓ 93K
Meta-Llama-3-8B logo
Meta-Llama-3-8B
NousResearch

Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the av…

Open Source 8.0B ↓ 64.5K
phi-1_5 logo
phi-1_5
microsoft

The language model Phi-1.5 is a Transformer with 1.3 billion parameters. It was trained using the same data sources as phi-1, augmented with a new data source that consists of various NLP synthetic texts. When assessed against benchmarks testing common sense, language understandi…

Open Source 1.3B ↓ 60.5K
Meta-Llama-3-70B-Instruct logo
Meta-Llama-3-70B-Instruct
meta-llama

Open Source 70.0B ↓ 59.5K
Phi-4-mini-reasoning logo
Phi-4-mini-reasoning
microsoft

Phi-4-mini-reasoning is a lightweight open model built upon synthetic data with a focus on high-quality, reasoning dense data further finetuned for more advanced math reasoning capabilities. The model belongs to the Phi-4 model family and supports 128K token context length.

Open Source ↓ 56.2K
granite-guardian-4.1-8b logo
granite-guardian-4.1-8b
ibm-granite

Granite Guardian 4.1 8B introduces improved Bring Your Own Criteria (BYOC) support, enabling users to define arbitrary judging criteria beyond the pre-baked safety and hallucination detectors. The model can now faithfully evaluate complex, multi-part requirements such as formatti…

Open Source 8.0B ↓ 55.4K
Meta-Llama-3-70B-Instruct logo
Meta-Llama-3-70B-Instruct
NousResearch

Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the av…

Open Source 70.0B ↓ 52.6K
Llama-3.2-1B logo
Llama-3.2-1B
NousResearch

The Meta Llama 3.2 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction-tuned generative models in 1B and 3B sizes (text in/text out). The Llama 3.2 instruction-tuned text only models are optimized for multilingual dialogue use cas…

Open Source 1.0B ↓ 48.9K
Hermes-2-Pro-Mistral-7B logo
Hermes-2-Pro-Mistral-7B
NousResearch

Hermes 2 Pro on Mistral 7B is the new flagship 7B Hermes!

Open Source 7.0B ↓ 48K
internlm2-7b logo
internlm2-7b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 25.2K