AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
MiniCPM-V-4_5-GPTQ logo
MiniCPM-V-4_5-GPTQ
openbmb

A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone

Open Source ↓ 462
internlm-20b logo
internlm-20b
internlm

The Shanghai Artificial Intelligence Laboratory, in collaboration with SenseTime Technology, the Chinese University of Hong Kong, and Fudan University, has officially released the 20 billion parameter pretrained model, InternLM-20B. InternLM-20B was pre-trained on over 2.3T Token…

Open Source 20.0B ↓ 450
japanese-stablelm-instruct-gamma-7b logo
japanese-stablelm-instruct-gamma-7b
stabilityai

This is a 7B-parameter decoder-only Japanese language model fine-tuned on instruction-following datasets, built on top of the base model Japanese Stable LM Base Gamma 7B.

Open Source 7.0B ↓ 440
Swallow-7b-hf logo
Swallow-7b-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 433
llama-3-youko-8b-instruct logo
llama-3-youko-8b-instruct
rinna

Llama 3 Youko 8B Instruct (rinna/llama-3-youko-8b-instruct)

Open Source 8.0B ↓ 402
japanese-stablelm-3b-4e1t-instruct logo
japanese-stablelm-3b-4e1t-instruct
stabilityai

This is a 3B-parameter decoder-only Japanese language model fine-tuned on instruction-following datasets, built on top of the base model Japanese StableLM-3B-4E1T Base.

Open Source 3.0B ↓ 389
Swallow-13b-instruct-hf logo
Swallow-13b-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 13.0B ↓ 380
japanese-stablelm-instruct-alpha-7b-v2 logo
japanese-stablelm-instruct-alpha-7b-v2
stabilityai

"A parrot able to speak Japanese, ukiyoe, edo period" — Stable Diffusion XL

Open Source 7.0B ↓ 377
Swallow-7b-plus-hf logo
Swallow-7b-plus-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 274
bilingual-gpt-neox-4b-instruction-sft logo
bilingual-gpt-neox-4b-instruction-sft
rinna

- 2023/08/02 We uploaded the newly trained rinna/bilingual-gpt-neox-4b-instruction-sft with the MIT license. - Please refrain from using the previous model released on 2023/07/31 for commercial purposes if you have already downloaded it. - The new model released on 2023/08/02 is…

Open Source 4.0B ↓ 268
internlm-chat-20b logo
internlm-chat-20b
internlm

The Shanghai Artificial Intelligence Laboratory, in collaboration with SenseTime Technology, the Chinese University of Hong Kong, and Fudan University, has officially released the 20 billion parameter pretrained model, InternLM-20B. InternLM-20B was pre-trained on over 2.3T Token…

Open Source 20.0B ↓ 249
internlm2-chat-1_8b-sft logo
internlm2-chat-1_8b-sft
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 8.0B ↓ 248