AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
MiniCPM-Llama3-V-2_5 logo
MiniCPM-Llama3-V-2_5
openbmb

A GPT-4V Level Multimodal LLM on Your Phone

Multimodal 8.5B ↓ 22.7K
internlm2-20b logo
internlm2-20b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 20.0B ↓ 21.3K
internlm2-base-7b logo
internlm2-base-7b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 21.2K
Llama-2-7b-chat-hf logo
Llama-2-7b-chat-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 7B fine-tuned model, optimized for dialogue use cases and converted for the Hugging Face Transformers forma…

Open Source 7.0B ↓ 17.9K
internlm2-base-20b logo
internlm2-base-20b
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 20.0B ↓ 17.3K
MiniCPM-V-4_5-AWQ logo
MiniCPM-V-4_5-AWQ
openbmb

A GPT-4o Level MLLM for Single Image, Multi Image and Video Understanding on Your Phone

Multimodal 8.7B ↓ 13.7K
stablelm-zephyr-3b logo
stablelm-zephyr-3b
stabilityai

Please note: For commercial use, please refer to https://stability.ai/license.

Open Source 3.0B ↓ 12K
Nous-Hermes-2-Mixtral-8x7B-DPO logo
Nous-Hermes-2-Mixtral-8x7B-DPO
NousResearch

Nous Hermes 2 Mixtral 8x7B DPO is the new flagship Nous Research model trained over the Mixtral 8x7B MoE LLM.

Open Source 7.0B ↓ 11.9K
Hermes-2-Theta-Llama-3-8B logo
Hermes-2-Theta-Llama-3-8B
NousResearch

Hermes-2 Θ (Theta) is the first experimental merged model released by Nous Research, in collaboration with Charles Goddard at Arcee, the team behind MergeKit.

Open Source 8.0B ↓ 11.8K
Infinity-Instruct-3M-0625-Llama3-8B logo
Infinity-Instruct-3M-0625-Llama3-8B
BAAI

Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][🤗] (would be released soon)

Open Source 8.0B ↓ 8.7K
Infinity-Instruct-3M-0625-Yi-1.5-9B logo
Infinity-Instruct-3M-0625-Yi-1.5-9B
BAAI

Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][🤗] (would be released soon)

Open Source 9.0B ↓ 8.7K
tulu-2-7b logo
tulu-2-7b
allenai

Tulu is a series of language models that are trained to act as helpful assistants. Tulu 2 7B is a fine-tuned version of Llama 2 that was trained on a mix of publicly available, synthetic and human datasets.

Open Source 7.0B ↓ 8.5K