AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
Infinity-Instruct-7M-Gen-Llama3_1-70B logo
Infinity-Instruct-7M-Gen-Llama3_1-70B
BAAI

Beijing Academy of Artificial Intelligence (BAAI) [Paper][Code][🤗] (would be released soon)

Open Source 70.0B ↓ 239
internlm2-chat-20b-sft logo
internlm2-chat-20b-sft
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 20.0B ↓ 234
Swallow-70b-instruct-hf logo
Swallow-70b-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 70.0B ↓ 216
bilingual-gpt-neox-4b-instruction-ppo logo
bilingual-gpt-neox-4b-instruction-ppo
rinna

Overview This repository provides an English-Japanese bilingual GPT-NeoX model of 3.8 billion parameters.

Open Source 4.0B ↓ 213
Meta-Llama-3-8B-Alternate-Tokenizer logo
Meta-Llama-3-8B-Alternate-Tokenizer
NousResearch

An alternate Meta-Llama-3-8B Repo for the Hermes Tokenizer

Open Source 8.0B ↓ 177
Swallow-70b-NVE-hf logo
Swallow-70b-NVE-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 70.0B ↓ 170
youri-7b-chat logo
youri-7b-chat
rinna

Overview The model is the instruction-tuned version of rinna/youri-7b . It adopts a chat-style input format.

Open Source 7.0B ↓ 152
Swallow-7b-NVE-hf logo
Swallow-7b-NVE-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 143
Swallow-7b-NVE-instruct-hf logo
Swallow-7b-NVE-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 141
Swallow-70b-hf logo
Swallow-70b-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 70.0B ↓ 134
Swallow-70b-NVE-instruct-hf logo
Swallow-70b-NVE-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 70.0B ↓ 134
Swallow-13b-NVE-hf logo
Swallow-13b-NVE-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 13.0B ↓ 120