LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
WeDLM-8B-Instruct is our flagship instruction-tuned diffusion language model that performs parallel decoding under standard causal attention, fine-tuned from WeDLM-8B.
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
This model is part of the GELab-Zero project, which presents the Step-GUI Technical Report [Paper] [Project Page] [Code].
This repository provides large language models trained by SB Intuitions.
State-of-the-art bilingual open-sourced Math reasoning LLMs. A solver , prover , verifier , augmentor .
Tiny Language Model that thinks in Japanese
The Shanghai Artificial Intelligence Laboratory, in collaboration with SenseTime Technology, the Chinese University of Hong Kong, and Fudan University, has officially released the 20 billion parameter pretrained model, InternLM-20B. InternLM-20B was pre-trained on over 2.3T Token…
Ling is a MoE LLM provided and open-sourced by InclusionAI. We introduce two different sizes, which are Ling-Lite and Ling-Plus. Ling-Lite has 16.8 billion parameters with 2.75 billion activated parameters, while Ling-Plus has 290 billion parameters with 28.8 billion activated pa…
We introduce EXAONE Deep, which exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research. Evaluation results show that 1) EXAONE Deep 2.4B outperforms other models…
🤗 Hugging Face 🖥️ Official Website 🕖 HunyuanAPI 🕹️ Demo 🤖 ModelScope
A cute robot wearing a kimono writes calligraphy with one single brush — Stable Diffusion XL
📃 License • 👨💻 Code • 🖥️ Demo • 📑 Technical Report • 📊 Benchmarks • 🚀 Getting Started