Back to model catalog
Compare models
Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.
| Attribute | Grok 4.5x-ai/grok-4.5 | Grok 4.3x-ai/grok-4.3 | Grok Build 0.1x-ai/grok-build-0.1 |
|---|---|---|---|
| Specs | |||
| Context window | 500K | 1M | 256K |
| Max output | 500K | 1M | 256K |
| Input modalities | TextImage | TextImage | TextImage |
| Output modalities | Text | Text | Text |
| Capabilities | |||
| Streaming | Supported | Supported | Supported |
| Tool calling | Supported | Supported | Supported |
| Structured output (JSON) | Supported | Supported | Supported |
| Image input | Supported | Supported | Supported |
| Thinking mode | Supported | Supported | Supported |
| Price / 1M tokens | |||
| Input | $ 2.00>= 200K$ 4.00 | $ 1.25>= 200K$ 2.50 | $ 1.00>= 200K$ 2.00 |
| Output | $ 6.00>= 200K$ 12.00 | $ 2.50>= 200K$ 5.00 | $ 2.00>= 200K$ 4.00 |
| Cache read | $ 0.30>= 200K$ 0.60 | $ 0.20>= 200K$ 0.40 | $ 0.20>= 200K$ 0.40 |
| Cache write | $ 2.00>= 200K$ 4.00 | $ 1.25>= 200K$ 2.50 | $ 1.00>= 200K$ 2.00 |
| What it is good at | |||
| Best for |
|
|
|
| Published benchmarks | |||
| SWE Marathonpass_at_1 |
| — | — |
| Terminal-Bench 2.1pass_ratesetups differ by: reasoning |
|
|
|
| DeepSWE 1.0resolved_rate |
| — | — |
| DeepSWE 1.1resolved_rate |
| — | — |
| SWE-Bench Proresolved_rate |
| — | — |
| SciCodescoresetups differ by: reasoning | — |
|
|
| GDPval-AA v2elosetups differ by: reasoning | — |
|
|
| AA-LCRscoresetups differ by: reasoning | — |
|
|
| AA Intelligence Index v4.1.1scoresetups differ by: reasoning | — |
|
|
| Humanity's Last Examaccuracysetups differ by: reasoning | — |
|
|
| CritPtscoresetups differ by: reasoning | — |
|
|
| GPQA Diamondaccuracysetups differ by: reasoning | — |
|
|
| tau3-Bankingscoresetups differ by: reasoning | — |
|
|