Back to model catalog

Compare models

Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.

DeepSeek V4 FlashDeepSeek V4 ProChat with these 2
AttributeDeepSeek V4 Flashdeepseek/deepseek-v4-flash10% offDeepSeek V4 Prodeepseek/deepseek-v4-pro10% off
Specs
Context window1M1M
Max output384K384K
Input modalitiesTextText
Output modalitiesTextText
Capabilities
StreamingSupportedSupported
Tool callingSupportedSupported
Structured output (JSON)SupportedSupported
Image inputNot supportedNot supported
Thinking modeSupportedSupported
Price / 1M tokens
Input$ 0.45$ 0.41$ 1.35$ 1.22
Output$ 1.35$ 1.22$ 4.05$ 3.65
Cache read$ 0.015$ 0.013$ 0.045$ 0.041
Cache write$ 0.45$ 0.41$ 1.35$ 1.22
What it is good at
Best for
  • Cost-sensitive coding agents
  • High-volume tool workflows
  • Complex long-horizon agents
  • Research and hard reasoning
Published benchmarks
Terminal-Bench 2.1pass_rate
  • 78.7%effort max(setup not fully disclosed)
  • 78.7%effort max(setup not fully disclosed)
SciCodescore
  • 49.9%effort max(setup not fully disclosed)
  • 49.2%effort max(setup not fully disclosed)
GDPval-AA v2elo
  • 1558.4effort max(setup not fully disclosed)
  • 1590.3effort max(setup not fully disclosed)
AA-LCRscore
  • 74.3%effort max(setup not fully disclosed)
  • 75.3%effort max(setup not fully disclosed)
AA Intelligence Index v4.1.1score
  • 51.8effort max(setup not fully disclosed)
  • 53effort max(setup not fully disclosed)
Humanity's Last Examaccuracy
  • 38.6%effort max(setup not fully disclosed)
  • 39.3%effort max(setup not fully disclosed)
CritPtscore
  • 16.6%effort max(setup not fully disclosed)
  • 18%effort max(setup not fully disclosed)
GPQA Diamondaccuracy
  • 90.8%effort max(setup not fully disclosed)
  • 92.8%effort max(setup not fully disclosed)
tau3-Bankingscore
  • 39.4%effort max(setup not fully disclosed)
  • 39.6%effort max(setup not fully disclosed)