All AI models · LLM benchmarks · Methodology

Z.AI

GLM 4.6 Turbo

Fast variant of GLM 4.6 for general chat, coding, and analysis with improved latency and strong reasoning.

Model facts

Context window
205K tokens
Maximum output
131K tokens
Cheapest paid input
$1 per 1M tokens
Cheapest paid output
$3 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

open weights, text.

Observed price history

  • 2026-08-24: $1 input / $3 output per 1M tokens

API providers