All AI models · LLM benchmarks · Methodology

OpenAI

GPT-5.6 Luna

Cost-efficient GPT-5.6 model for fast, high-volume workloads

Model facts

Context window
1.05M tokens
Maximum output
128K tokens
Cheapest paid input
$0.1 per 1M tokens
Cheapest paid output
$0.6 per 1M tokens
Open weights
no
Providers
45

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, text, image, pdf.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
max52.371.4not measured116.2
xhigh50.168.6not measured113.1
high47.063.3not measured106.4
medium38.950.7not measured106.2
low33.944.2not measured111.6
Non-reasoning26.839.3not measured107.6

Observed price history

  • 2026-08-24: $0.1 input / $0.6 output per 1M tokens
  • 2026-08-28: $0.2 input / $1.2 output per 1M tokens
  • 2026-08-31: $0.1 input / $0.6 output per 1M tokens

API providers