AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

63 models for "Open Weights" Compare
OLMo-2-1124-7B-SFT logo
OLMo-2-1124-7B-SFT
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 6.2K
OLMo-2-0325-32B logo
OLMo-2-0325-32B
allenai

We introduce OLMo 2 32B, the largest model in the OLMo 2 family. OLMo 2 was pre-trained on OLMo-mix-1124 and uses Dolmino-mix-1124 for mid-training.

Open Source 32.0B ↓ 5.7K
OLMo-2-1124-7B-DPO logo
OLMo-2-1124-7B-DPO
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 3.7K
Molmo-7B-O-0924 logo
Molmo-7B-O-0924
allenai

Molmo is a family of open vision-language models developed by the Allen Institute for AI. Molmo models are trained on PixMo, a dataset of 1 million, highly-curated image-text pairs. It has state-of-the-art performance among multimodal models with a similar size while being fully…

Multimodal 7.7B ↓ 3.5K
Meta-Llama-Guard-2-8B logo
Meta-Llama-Guard-2-8B
meta-llama

Open Source 8.0B ↓ 2.2K
Llama-Guard-3-11B-Vision logo
Llama-Guard-3-11B-Vision
meta-llama

Multimodal 11.0B ↓ 2.1K
GroveMoE-Inst logo
GroveMoE-Inst
inclusionAI

GroveMoE-Inst 🤗 Models &nbsp&nbsp &nbsp&nbsp 📑 Paper &nbsp&nbsp &nbsp&nbsp 🔗 Github &nbsp&nbsp

Open Source 33.0B ↓ 1.8K
Intern-S2-397B logo
Intern-S2-397B
internlm

💻Github Repo • 🤗HF Model Collections • 🤖ModelScope Collections • 💬Online Chat

Multimodal 397.0B ↓ 1.1K
deepseek-coder-5.7bmqa-base logo
deepseek-coder-5.7bmqa-base
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 5.7B ↓ 918
LLaDA-UI logo
LLaDA-UI
inclusionAI

Bringing Block-wise Diffusion to Vision-Language GUI Agents

Multimodal 16.7B ↓ 760
SmolLM3-3B-ONNX logo
SmolLM3-3B-ONNX
HuggingFaceTB

1. Model Summary 2. How to use 3. Evaluation 4. Training 5. Limitations 6. License

Open Source 3.0B ↓ 442
Intern-S2-397B-FP8 logo
Intern-S2-397B-FP8
internlm

💻Github Repo • 🤗HF Model Collections • 🤖ModelScope Collections • 💬Online Chat

Multimodal 397.0B ↓ 310