All AI models · LLM benchmarks · Methodology

Z.AI

GLM 4.6 Turbo (Thinking)

GLM 4.6 Turbo with thinking mode enabled for enhanced reasoning; shows internal reasoning and supports long context.

Model facts

Context window
205K tokens
Maximum output
131K tokens
Cheapest paid input
$1 per 1M tokens
Cheapest paid output
$3 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

reasoning, open weights, text.

Observed price history

  • 2026-08-24: $1 input / $3 output per 1M tokens

API providers