All AI models · LLM benchmarks · Methodology

DeepSeek

DeepSeek V4 Pro 0813

DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes

Model facts

Context window
1.05M tokens
Maximum output
384K tokens
Cheapest paid input
$0.35 per 1M tokens
Cheapest paid output
$0.8 per 1M tokens
Open weights
yes
Providers
38

Capabilities and modalities

reasoning, tool calling, structured output, temperature control, open weights, text.

Local hardware estimate

At 4K context: Q4_K_M 1011.2 GiB, Q8_0 1779.7 GiB, F16 3220.8 GiB working memory.

Parameters
1650.5 billion
Architecture
deepseekv4
Model license
mit

Planning estimate adapted from the MIT-licensed whichllm estimator using metadata from Hugging Face; not a vendor minimum requirement.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
Reasoning, Max Effort53.268.8not measured66.9

Observed price history

  • 2026-08-24: $0.6 input / $1.9 output per 1M tokens
  • 2026-08-25: $0.581 input / $1.74 output per 1M tokens
  • 2026-08-26: $0.581 input / $1.74 output per 1M tokens
  • 2026-08-28: $0.35 input / $0.8 output per 1M tokens

API providers

  • Alibaba Token Plan — model id deepseek-v4-pro-0813; free input / free output per 1M tokens; provider documentation
  • Alibaba Token Plan (China) — model id deepseek-v4-pro-0813; free input / free output per 1M tokens; provider documentation
  • Nvidia — model id deepseek-ai/deepseek-v4-pro-0813; free input / free output per 1M tokens; provider documentation
  • SCNet Token Plan — model id DeepSeek-V4-Pro-0813; free input / free output per 1M tokens; provider documentation
  • CrofAI — model id deepseek-v4-pro-0813; $0.35 input / $0.8 output per 1M tokens; provider documentation
  • OrcaRouter — model id deepseek/deepseek-v4-pro-0813; $0.442 input / $0.884 output per 1M tokens; provider documentation
  • Eden AI — model id qwen/deepseek-v4-pro-0813; $0.581 input / $1.74 output per 1M tokens; provider documentation
  • RunInfra — model id deepseek-ai/DeepSeek-V4-Pro-0813; $0.6 input / $1.9 output per 1M tokens; provider documentation
  • Merge Gateway — model id deepseek/deepseek-v4-pro-0813; $0.66 input / $1.98 output per 1M tokens; provider documentation
  • Vercel AI Gateway — model id deepseek/deepseek-v4-pro-0813; $0.66 input / $1.98 output per 1M tokens; provider documentation
  • AIHubMix — model id deepseek-v4-pro-0813; $0.692 input / $2.08 output per 1M tokens; provider documentation
  • NanoGPT — model id deepseek/deepseek-v4-pro-0813; $1.1 input / $2.5 output per 1M tokens; provider documentation
  • Deep Infra — model id deepseek-ai/DeepSeek-V4-Pro-0813; $1.3 input / $2.6 output per 1M tokens; provider documentation
  • Eden AI — model id deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813; $1.3 input / $2.6 output per 1M tokens; provider documentation
  • OpenRouter — model id deepseek/deepseek-v4-pro-0813; $1.12 input / $3.35 output per 1M tokens; provider documentation
  • Requesty (EU) — model id deepseek-v4-pro-0813@eu; $1.75 input / $3.5 output per 1M tokens; provider documentation
  • Weights & Biases — model id deepseek-ai/DeepSeek-V4-Pro-0813; $1.31 input / $3.96 output per 1M tokens; provider documentation
  • Eden AI — model id databricks/databricks-deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Arcee — model id deepseek/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Baseten — model id deepseek-ai/DeepSeek-V4-Pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Cloudflare Workers AI — model id @cf/deepseek-ai/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • DigitalOcean — model id deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Eden AI — model id cloudflare/@cf/deepseek-ai/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Eden AI — model id fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Eden AI — model id together_ai/deepseek-ai/DeepSeek-V4-Pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • EmpirioLabs AI — model id deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Fireworks AI — model id accounts/fireworks/models/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Hugging Face — model id deepseek-ai/DeepSeek-V4-Pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Kilo Gateway — model id deepseek/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Ofox — model id deepseek/deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • OpenRouter (batch) — model id deepseek/deepseek-v4-pro-0813:batch; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Requesty — model id deepseek-v4-pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Together AI — model id deepseek-ai/DeepSeek-V4-Pro-0813; $1.32 input / $3.96 output per 1M tokens; provider documentation
  • Volcengine Ark — model id deepseek-v4-pro-ga-260813; $1.34 input / $4.01 output per 1M tokens; provider documentation
  • Charm Hyper — model id deepseek-v4-pro-0813; $1.44 input / $4.31 output per 1M tokens; provider documentation
  • Cortecs — model id deepseek-v4-pro-0813; $2 input / $4 output per 1M tokens; provider documentation
  • Eden AI — model id tensorx/deepseek/deepseek-v4-pro-0813; $2 input / $4 output per 1M tokens; provider documentation
  • Venice AI — model id deepseek-v4-pro-0813; $1.65 input / $4.95 output per 1M tokens; provider documentation