All AI models · LLM benchmarks · Methodology

Z.AI

GLM 4.6V Original

GLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).

Model facts

Context window
128K tokens
Maximum output
24K tokens
Cheapest paid input
$0.6 per 1M tokens
Cheapest paid output
$0.9 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

attachments, open weights, text, image.

Observed price history

  • 2026-08-24: $0.6 input / $0.9 output per 1M tokens

API providers