AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
Llama-2-70b-chat-hf logo
Llama-2-70b-chat-hf
meta-llama

Open Source 70.0B ↓ 8K
granite-guardian-3.2-3b-a800m logo
granite-guardian-3.2-3b-a800m
ibm-granite

Granite Guardian 3.2 3B-A800M is a fine-tuned Granite 3.2 3B-A800M instruct model designed to detect risks in prompts and responses. It can help with risk detection along many key dimensions catalogued in the IBM AI Risk Atlas. It is trained on unique data comprising human annota…

Open Source 3.3B ↓ 7.3K
internlm2-1_8b logo
internlm2-1_8b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 8.0B ↓ 6.7K
Falcon3-1B-Base logo
Falcon3-1B-Base
tiiuae

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B parameters.

Open Source 1.0B ↓ 6.1K
internlm2-chat-1_8b logo
internlm2-chat-1_8b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 8.0B ↓ 5.8K
granite-guardian-3.1-2b logo
granite-guardian-3.1-2b
ibm-granite

Granite Guardian 3.1 2B is a fine-tuned Granite 3.1 2B Instruct model designed to detect risks in prompts and responses. It can help with risk detection along many key dimensions catalogued in the IBM AI Risk Atlas. It is trained on unique data comprising human annotations and sy…

Open Source 2.0B ↓ 5K
Llama-2-13b-hf logo
Llama-2-13b-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 13B pretrained model, converted for the Hugging Face Transformers format. Links to other models can be foun…

Open Source 13.0B ↓ 4.8K
internlm2-chat-7b-sft logo
internlm2-chat-7b-sft
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 4.4K
Hermes-2-Pro-Llama-3-8B logo
Hermes-2-Pro-Llama-3-8B
NousResearch

Hermes 2 Pro is an upgraded, retrained version of Nous Hermes 2, consisting of an updated and cleaned version of the OpenHermes 2.5 Dataset, as well as a newly introduced Function Calling and JSON Mode dataset developed in-house.

Open Source 8.0B ↓ 4K
Hermes-2-Theta-Llama-3-70B logo
Hermes-2-Theta-Llama-3-70B
NousResearch

Hermes-2 Θ (Theta) 70B is the continuation of our experimental merged model released by Nous Research, in collaboration with Charles Goddard and Arcee AI, the team behind MergeKit.

Open Source 70.0B ↓ 3.7K
Llama-2-70b-hf logo
Llama-2-70b-hf
meta-llama

Open Source 70.0B ↓ 3.4K
MiniCPM-V-4_5-int4 logo
MiniCPM-V-4_5-int4
openbmb

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

Multimodal 8.7B ↓ 3.3K