All AI models · LLM benchmarks · Methodology

OpenAI

GPT-4 Turbo

Compact GPT model for low-latency assistance and high-volume workloads

Model facts

Context window
128K tokens
Maximum output
4K tokens
Cheapest paid input
$5 per 1M tokens
Cheapest paid output
$15 per 1M tokens
Open weights
no
Providers
16

Capabilities and modalities

tool calling, structured output, attachments, temperature control, text, image, pdf.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
default7.721.5not measurednot measured

Observed price history

  • 2026-08-24: $5 input / $15 output per 1M tokens
  • 2026-08-28: $9 input / $27 output per 1M tokens
  • 2026-08-31: $5 input / $15 output per 1M tokens

API providers