LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
💻Github Repo • 🤗Model Collections • 📜Technical Report
1. Model Summary 2. Use 3. Training 4. Citation
Redmond-Hermes-Coder 15B is a state-of-the-art language model fine-tuned on over 300,000 instructions. This model was fine-tuned by Nous Research, with Teknium and Karan4D leading the fine tuning process and dataset curation, Redmond AI sponsoring the compute, and several other c…
Qwen-SEA-LION-v4-32B-IT-8BIT (GPTQ model)
SEA-Safeguard is a collection of safety-focused Large Language Models (LLMs) built upon the SEA-LION family, designed specifically for the Southeast Asia (SEA) region.
One of the focus areas at Together Research is new architectures for long context, improved training, and inference performance over the Transformer architecture. Spinning out of a research program from our team and academic collaborators, with roots in signal processing-inspired…
💻Github Repo • 🤗Model Collections • 📜Technical Report
💻Github Repo • 🤗Model Collections • 📜Technical Report • 💬Online Chat
A cute robot wearing a kimono writes calligraphy with one single brush — Stable Diffusion XL
Qwen-SEA-LION-v4-32B-IT-4BIT (GPTQ model)
"A parrot able to speak Japanese, ukiyoe, edo period" — Stable Diffusion XL
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]