All AI models · LLM benchmarks · Methodology

DeepSeek

DeepSeek V4 Flash Latest

This model always redirects to the latest model in the DeepSeek V4 Flash family.

Model facts

Context window
1.05M tokens
Maximum output
393K tokens
Cheapest paid input
$0.05 per 1M tokens
Cheapest paid output
$0.16 per 1M tokens
Open weights
yes
Providers
4

Capabilities and modalities

reasoning, tool calling, structured output, temperature control, open weights, text.

Observed price history

  • 2026-08-24: $0.035 input / $0.12 output per 1M tokens
  • 2026-08-25: $0.035 input / $0.1 output per 1M tokens
  • 2026-08-26: $0.03 input / $0.075 output per 1M tokens
  • 2026-08-27: $0.03 input / $0.1 output per 1M tokens
  • 2026-08-30: $0.03 input / $0.16 output per 1M tokens
  • 2026-09-01: $0.05 input / $0.1 output per 1M tokens
  • 2026-09-02: $0.05 input / $0.16 output per 1M tokens

API providers

  • OpenRouter — model id ~deepseek/deepseek-v4-flash-latest; $0.05 input / $0.16 output per 1M tokens; provider documentation
  • Kilo Gateway — model id ~deepseek/deepseek-v4-flash-latest; $0.05 input / $0.16 output per 1M tokens; provider documentation
  • Arcee — model id deepseek/deepseek-v4-flash-latest; $0.14 input / $0.28 output per 1M tokens; provider documentation
  • NanoGPT — model id deepseek/deepseek-v4-flash-latest; $0.14 input / $0.28 output per 1M tokens; provider documentation