LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
This model is part of the GELab-Zero project, which presents the Step-GUI Technical Report [Paper] [Project Page] [Code].
Tiny Language Model that thinks in Japanese
The Shanghai Artificial Intelligence Laboratory, in collaboration with SenseTime Technology, the Chinese University of Hong Kong, and Fudan University, has officially released the 20 billion parameter pretrained model, InternLM-20B. InternLM-20B was pre-trained on over 2.3T Token…
Ling is a MoE LLM provided and open-sourced by InclusionAI. We introduce two different sizes, which are Ling-Lite and Ling-Plus. Ling-Lite has 16.8 billion parameters with 2.75 billion activated parameters, while Ling-Plus has 290 billion parameters with 28.8 billion activated pa…
We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research. Evaluation results show that 1) EXAONE Deep 2.4B outperforms other models…
🤗 Hugging Face 🖥️ Official Website 🕖 HunyuanAPI 🕹️ Demo 🤖 ModelScope
This repository is primarily for the Transformers framework. If you're using other open-source frameworks, please use the alternative repository: MiniMax-M1-40k
The first commercially available language model released by Nous Research!
💻Github Repo • 🤗Model Collections • 🌳Arch Space
[AgentOhana Paper] [Github] [Discord] [Homepage] [Community Demo]
AquilaMoE: Efficient Training for MoE Models with Scale-Up and Scale-Out Strategies Language Foundation Model & Software Team Beijing Academy of Artificial Intelligence (BAAI) [Paper(released soon)] [Code] [github]
[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.