All AI models · LLM benchmarks · Methodology

Z.AI

GLM-5.2

Open flagship GLM for long-horizon coding agents and million-token context work

Model facts

Context window
1M tokens
Maximum output
131K tokens
Cheapest paid input
$0.3 per 1M tokens
Cheapest paid output
$1.05 per 1M tokens
Open weights
yes
Providers
101

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, open weights, text, image.

Local hardware estimate

At 4K context: Q4_K_M 462.1 GiB, Q8_0 812.8 GiB, F16 1470.6 GiB working memory.

Parameters
753.3 billion
Architecture
glmmoedsa
Model license
mit

Planning estimate adapted from the MIT-licensed whichllm estimator using metadata from Hugging Face; not a vendor minimum requirement.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
max52.668.8not measured68.6
Non-reasoning34.846.5not measured93.4

Observed price history

  • 2026-08-24: $0.3 input / $1.05 output per 1M tokens

API providers