Back to model catalog

Compare models

Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.

DeepSeek V4.1 FlashChat with this model
AttributeDeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash10% off
Specs
Context window1M
Max output384K
Input modalitiesTextImage
Output modalitiesText
Capabilities
StreamingSupported
Tool callingSupported
Structured output (JSON)Supported
Image inputSupported
Thinking modeSupported
Price / 1M tokens
Input$ 0.30$ 0.27
Output$ 1.20$ 1.08
Cache read$ 0.0060$ 0.0054
Cache write$ 0.30$ 0.27
What it is good at
Best for
  • Terminal and coding agents
  • Agents that read screenshots and charts
  • Hard knowledge and research
  • High-throughput production traffic
Published benchmarks
Terminal-Bench 2.1pass_rate
  • 90.6%(setup not fully disclosed)
Terminal-Bench 4.0pass_rate
  • 31.2%(setup not fully disclosed)
Agents' Last Examscore
  • 31.8%(setup not fully disclosed)
AutomationBenchscore
  • 54.8%(setup not fully disclosed)
DeepSWE 1.1resolved_rate
  • 74.2%(setup not fully disclosed)
NL2Reposcore
  • 65.4%(setup not fully disclosed)
CyberGymscore
  • 88.1%(setup not fully disclosed)
MathArena Apexscore
  • 65.6%(setup not fully disclosed)
Chartography (with tools)score
  • 78.9%(setup not fully disclosed)
Humanity's Last Examaccuracy
  • 36.8%(setup not fully disclosed)
Humanity's Last Exam (with tools)accuracy
  • 63.9%(setup not fully disclosed)
GPQA Diamondaccuracy
  • 90.9%(setup not fully disclosed)