Back to model catalog

Compare models

Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.

GPT-4o MiniGPT-5.4 NanoChat with these 2
AttributeGPT-4o Miniopenai/gpt-4o-miniGPT-5.4 Nanoopenai/gpt-5.4-nano
Specs
Context window128K400K
Max output16K128K
Input modalitiesTextImageTextImage
Output modalitiesTextText
Capabilities
StreamingSupportedSupported
Tool callingSupportedSupported
Structured output (JSON)SupportedSupported
Image inputSupportedSupported
Thinking modeNot supportedSupported
Price / 1M tokens
Input$ 0.15$ 0.20
Output$ 0.60$ 1.25
Cache read$ 0.075$ 0.020
Cache write$ 0.15$ 0.20
What it is good at
Best for
  • High-volume focused tasks
  • Translation and text transformation
  • Classification and extraction
  • Ranking and simple subagents
Published benchmarks
Terminal-Bench 2.0pass_rate
  • 46.3%effort xhigh(setup not fully disclosed)
HumanEvalpass_at_1
  • 87.2%(setup not fully disclosed)
SWE-Bench Pro (Public)pass_rate
  • 52.4%effort xhigh(setup not fully disclosed)
OSWorld-Verifiedsuccess_rate
  • 39%effort xhigh(setup not fully disclosed)
OpenAI MRCR v2 8-needle 128K–256Kaccuracy
  • 33.1%effort xhigh(setup not fully disclosed)
MGSMaccuracy
  • 87%(setup not fully disclosed)
MMMU Pro (no tools)accuracy
  • 66.1%effort xhigh(setup not fully disclosed)
Humanity's Last Exam Full (with tools)accuracy
  • 37.7%effort xhigh(setup not fully disclosed)
MMLUaccuracy
  • 82%(setup not fully disclosed)
GPQA Diamondaccuracy
  • 82.8%effort xhigh(setup not fully disclosed)
Toolathlonscore
  • 35.5%effort xhigh(setup not fully disclosed)