Back to model catalog

Compare models

Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.

Gemini 2.5 ProGemini 2.5 FlashChat with these 2
AttributeGemini 2.5 Progemini/gemini-2.5-proGemini 2.5 Flashgemini/gemini-2.5-flash
Specs
Context window1M1M
Max output66K66K
Input modalitiesTextImageAudioVideoTextImageAudioVideo
Output modalitiesTextText
Capabilities
StreamingSupportedSupported
Tool callingSupportedSupported
Structured output (JSON)SupportedSupported
Image inputSupportedSupported
Thinking modeSupportedSupported
Price / 1M tokens
Input$ 1.25>= 200K$ 2.50$ 0.30
Output$ 10.00>= 200K$ 15.00$ 2.50
Cache read$ 0.13>= 200K$ 0.25$ 0.030
Cache write$ 1.25>= 200K$ 2.50$ 0.30
What it is good at
Best for
  • Long-context analysis
  • Math and STEM reasoning
  • High-volume reasoning
  • Multimodal analysis
Published benchmarks
SWE-Bench Verifiedresolved_ratesetups differ by: harness, reasoning
  • 63.8%custom-agent(setup not fully disclosed)
  • 60.4%enabled(setup not fully disclosed)
AIME 2025accuracy
  • 88%enabled(setup not fully disclosed)
  • 72%enabled(setup not fully disclosed)
MMMUaccuracysetups differ by: reasoning
  • 84%(setup not fully disclosed)
  • 79.7%enabled(setup not fully disclosed)
Humanity's Last Examaccuracysetups differ by: reasoning
  • 18.8%(setup not fully disclosed)
  • 11%enabled(setup not fully disclosed)
GPQA Diamondaccuracy
  • 86.4%enabled(setup not fully disclosed)
  • 82.8%enabled(setup not fully disclosed)