All AI models · LLM benchmarks · Methodology

Z.AI

GLM 5.3 Flash Uncensored

GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.

Model facts

Context window
1.05M tokens
Maximum output
33K tokens
Cheapest paid input
$0.35 per 1M tokens
Cheapest paid output
$1.4 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

reasoning, tool calling, structured output, attachments, open weights, text, image.

Observed price history

  • 2026-08-27: $0.125 input / $0.5 output per 1M tokens
  • 2026-08-30: $0.35 input / $1.4 output per 1M tokens

API providers

  • NanoGPT — model id z-ai/glm-5.3-flash-uncensored; $0.35 input / $1.4 output per 1M tokens; provider documentation