LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Stable LM 2 12B Chat is a 12 billion parameter instruction tuned language model trained on a mix of publicly available datasets and synthetic datasets, utilizing Direct Preference Optimization (DPO).
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
Use Stable Chat (Research Preview) to test Stability AI's best language models for free
We opensource our Aquila2 series, now including Aquila2 , the base language models, namely Aquila2-7B and Aquila2-34B , as well as AquilaChat2 , the chat models, namely AquilaChat2-7B and AquilaChat2-34B , as well as the long-text chat models, namely AquilaChat2-7B-16k and Aquila…
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
Model description xGen-MM is a series of the latest foundational Large Multimodal Models (LMMs) developed by Salesforce AI Research. This series advances upon the successful designs of the BLIP series, incorporating fundamental enhancements that ensure a more robust and superior…
Model Summary This is a 1.8B model trained on Cosmopedia synthetic dataset.
Building the Next Generation of Open-Source and Bilingual LLMs
OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.
This model is based on the principles described in the paper Large Language Diffusion Models.
💻Github Repo • 🤔Reporting Issues • 📜Technical Report
Llama 3.3 Swallow is a large language model (70B) that was built by continual pre-training on the Meta Llama 3.3 model. Llama 3.3 Swallow enhanced the Japanese language capabilities of the original Llama 3.3 while retaining the English language capabilities. We use approximately…