AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

133 models for "RLHF" Compare
Meta-Llama-3.1-70B-Instruct logo
Meta-Llama-3.1-70B-Instruct
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 70.0B ↓ 2.9K
Llama-2-13b-chat-hf logo
Llama-2-13b-chat-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 13B fine-tuned model, optimized for dialogue use cases and converted for the Hugging Face Transformers form…

Open Source 13.0B ↓ 2.2K
AprielGuard logo
AprielGuard
ServiceNow-AI

1. Summary 2. Taxonomy 2. Evaluation 3. Training Details 4. How to Use 5. Intended Use 6. Limitations 7. License 8. Citation

Open Source 8.0B ↓ 1.9K
Falcon3-3B-Base logo
Falcon3-3B-Base
tiiuae

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B parameters.

Open Source 3.0B ↓ 1.9K
Llama-2-70b-chat-hf logo
Llama-2-70b-chat-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 70B fine-tuned model, optimized for dialogue use cases and converted for the Hugging Face Transformers form…

Open Source 70.0B ↓ 1.9K
Swallow-7b-instruct-hf logo
Swallow-7b-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 1.9K
Meta-Llama-3.1-70B logo
Meta-Llama-3.1-70B
NousResearch

The Meta Llama 3.1 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction tuned generative models in 8B, 70B and 405B sizes (text in/text out). The Llama 3.1 instruction tuned text only models (8B, 70B, 405B) are optimized for multil…

Open Source 70.0B ↓ 1.6K
stablelm-tuned-alpha-3b logo
stablelm-tuned-alpha-3b
stabilityai

StableLM-Tuned-Alpha is a suite of 3B and 7B parameter decoder-only language models built on top of the StableLM-Base-Alpha models and further fine-tuned on various chat and instruction-following datasets.

Open Source 3.0B ↓ 1.5K
Falcon3-10B-Base logo
Falcon3-10B-Base
tiiuae

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B parameters.

Open Source 10.0B ↓ 1.3K
Meta-Llama-3-70B logo
Meta-Llama-3-70B
NousResearch

Meta developed and released the Meta Llama 3 family of large language models (LLMs), a collection of pretrained and instruction tuned generative text models in 8 and 70B sizes. The Llama 3 instruction tuned models are optimized for dialogue use cases and outperform many of the av…

Open Source 70.6B ↓ 1.1K
stablelm-tuned-alpha-7b logo
stablelm-tuned-alpha-7b
stabilityai

StableLM-Tuned-Alpha is a suite of 3B and 7B parameter decoder-only language models built on top of the StableLM-Base-Alpha models and further fine-tuned on various chat and instruction-following datasets.

Open Source 7.0B ↓ 1.1K
Llama-2-7B-32K-Instruct logo
Llama-2-7B-32K-Instruct
togethercomputer

Llama-2-7B-32K-Instruct is an open-source, long-context chat model finetuned from Llama-2-7B-32K, over high-quality instruction and chat data. We built Llama-2-7B-32K-Instruct with less than 200 lines of Python script using Together API, and we also make the recipe fully availabl…

Open Source 7.0B ↓ 905