AI Agent Hub
Back to models
🤖

Z.ai: GLM 4.6V

8.0 / 10 131.1K context Proprietary

About this model

GLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media. It supports up to 128K tokens, processes complex page layouts...

Benchmark Scores

GSM8K
87.6
HumanEval
72.0

Technical Specs

  • Architecture: text+image+video->text
  • Context Window: 131,072 tokens
  • Input Modalities: text

Hardware Requirements

  • API-only (no local hardware needed)

Pricing

Input Output Currency
0.3 / 1M tokens 0.8999999999999999 / 1M tokens USD