All AI models · LLM benchmarks · Methodology

Kimi K3 Flex

Kimi K3 on the flex tier: discounted, best-effort latency, requests may be held under load

Model facts

Context window
1.05M tokens
Maximum output
1.05M tokens
Cheapest paid input
$1.95 per 1M tokens
Cheapest paid output
$9.75 per 1M tokens
Open weights
yes
Providers
1

Capabilities and modalities

reasoning, tool calling, structured output, attachments, open weights, text, image.

Observed price history

  • 2026-08-28: $1.95 input / $9.75 output per 1M tokens

API providers