LLM Models · inclusionAI
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
A DSpark speculator for Ling3. DSpark extends DFlash with target-model auxiliary features and a confidence head that dynamically chooses the number of draft tokens. The model was trained with SpecForge and is served with SGLang.
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
🤗 Hugging Face 🤖 ModelScope 🐙 ling.tbox.cn Tech Report
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter
This model is presented in the paper Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model.
LLaDA2.0-mini is a diffusion language model featuring a 16BA1B Mixture-of-Experts (MoE) architecture. As an enhanced, instruction-tuned iteration of the LLaDA series, it is optimized for practical applications.
🚀 LLaDA2.1-flash is now live on ZenmuxAI ! Try it via API 🛠️ or Chat 💬: https://zenmux.ai/inclusionai/llada2.1-flash
UI-Venus-1.5 model This repository contains the UI-Venus model from the report UI-Venus-1.5 Technical Report. UI-Venus 1.5 is a unified, end-to-end GUI Agent designed for robust real-world applications. The model family includes two dense (2B/8B) and one MoE (30B-A3B) variants to…