All AI models · LLM benchmarks · Methodology

Z.AI

GLM Flash Latest

This model always redirects to the latest model in the GLM Flash family.

Model facts

Context window
1.31M tokens
Maximum output
944K tokens
Cheapest paid input
$0.075 per 1M tokens
Cheapest paid output
$0.25 per 1M tokens
Open weights
no
Providers
2

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, text, image, video.

Observed price history

  • 2026-09-02: $0.075 input / $0.25 output per 1M tokens

API providers

  • Kilo Gateway (Z.ai) — model id ~z-ai/glm-flash-latest; $0.075 input / $0.25 output per 1M tokens; provider documentation
  • OpenRouter — model id ~z-ai/glm-flash-latest; $0.075 input / $0.25 output per 1M tokens; provider documentation