AI Agent Hub
Back to models
🤖

Z.ai: GLM 4.5V

8.0 / 10 65.5K context Proprietary

About this model

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...

Benchmark Scores

GSM8K
87.6
HumanEval
72.0

Technical Specs

  • Architecture: text+image->text
  • Context Window: 65,536 tokens
  • Input Modalities: text

Hardware Requirements

  • API-only (no local hardware needed)

Pricing

Input Output Currency
0.6 / 1M tokens 1.7999999999999998 / 1M tokens USD