LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
🤗 HuggingFace 📔 Technical Report 📰 Blog Play around! 🗨️ Xiaomi MiMo Studio 🎨 Xiaomi MiMo API Platform
🤗 HuggingFace 📔 Technical Report 📰 Blog Play around! 🗨️ Xiaomi MiMo Studio 🎨 Xiaomi MiMo API Platform
NVIDIA-Nemotron-3.5-Lightning-30B-A3B-Base-BF16
Base Model OpenRouter Announcement
Base Model OpenRouter Announcement
Description: The NVIDIA Qwen3.5-397B-A17B NVFP4 model is the quantized version of Alibaba's Qwen3.5-397B-A17B model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Qwen3.5-397B-A17B N…
We introduce Olmo 3, a new family of 7B and 32B models. This suite includes Base, Instruct, and Think variants. The Base models were trained using a staged training approach.
We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.
We introduce Olmo 3, a new family of 7B and 32B models. This suite includes Base, Instruct, and Think variants. The Base models were trained using a staged training approach.
We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.
🤗 Hugging Face 🤖 ModelScope 🐙 OpenRouter