LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
This is UltraLM-65b delta weights, a chat language model trained upon UltraChat
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
Based off Puffin 13B which was the first commercially available language model released by Nous Research!
Feel free to try out our OpenChatKit feedback app!
Developed by : Upstage Backbone Model : LLaMA Variations : It has different model parameter sizes and sequence lengths: 30B/1024, 30B/2048, 65B/1024 Language(s) : English Library : HuggingFace Transformers License : This model is under a Non-commercial Bespoke License and governe…
MUCH BETTER MISTRAL BASED VERSION IS OUT NOW AS CAPYBARA V1.9
Stable Beluga 1 is a Llama65B model fine-tuned on an Orca style Dataset
Developed by : Upstage Backbone Model : LLaMA Variations : It has different model parameter sizes and sequence lengths: 30B/1024, 30B/2048, 65B/1024 Language(s) : English Library : HuggingFace Transformers License : This model is under a Non-commercial Bespoke License and governe…
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.
Feel free to try out our OpenChatKit feedback app!
SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The sizes of the models range from 3 billion to 7 billion parameters.
WeDLM-8B-Instruct is our flagship instruction-tuned diffusion language model that performs parallel decoding under standard causal attention, fine-tuned from WeDLM-8B.