LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Developed by : Upstage Backbone Model : LLaMA-2 Language(s) : English Library : HuggingFace Transformers License : Fine-tuned checkpoints is licensed under the Non-Commercial Creative Commons license (CC BY-NC-4.0) Where to send comments : Instructions on how to provide feedback…
Llama 3 Youko 8B Instruct (rinna/llama-3-youko-8b-instruct)
This is a 3B-parameter decoder-only Japanese language model fine-tuned on instruction-following datasets, built on top of the base model Japanese StableLM-3B-4E1T Base.
[Homepage] [Paper] [Discord] [Dataset] [Github]
"A parrot able to speak Japanese, ukiyoe, edo period" — Stable Diffusion XL
--- inference: false language: - en tags: - instruction-finetuning pretty name: JudgeLM-100K task categories: - text-generation ---
Fine-tuned version of HuggingFaceTB/SmolLM3-3B-Base optimized for grade school math (GSM8K benchmark).
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]
--- inference: false language: - en tags: - instruction-finetuning pretty name: JudgeLM-100K task categories: - text-generation ---
Compute provided by PygmalionAI, thank you! Follow PygmalionAI on Twitter @pygmalion ai.
The Capybara series is the first Nous collection of dataset and models made by fine-tuning mostly on data created by Nous in-house.
Building the Next Generation of Open-Source and Bilingual LLMs