All AI models · LLM benchmarks · Methodology

DeepSeek

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash supports text and image input, reasoning, tool calling, and structured output with a 1M-token context window. This is a rate-limited beta with limited capacity, intended for testing rather than production use. Assume prompts and responses are logged by the provider and may be used for model training or service improvement. Do not send sensitive or confidential data.

Model facts

Context window
1M tokens
Maximum output
384K tokens
Cheapest paid input
$0.156 per 1M tokens
Cheapest paid output
$0.312 per 1M tokens
Open weights
no
Providers
1

Capabilities and modalities

reasoning, tool calling, structured output, attachments, text, image.

Observed price history

  • 2026-09-08: $0.156 input / $0.312 output per 1M tokens

API providers

  • NanoGPT — model id deepseek/deepseek-v4.1-flash; $0.156 input / $0.312 output per 1M tokens; provider documentation