Back to model catalog
Compare models
Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.
| Attribute | GLM-5.3-Flashz-ai/glm-5.3-flash50% off | GLM-5.3z-ai/glm-5.3 | GLM-5.2z-ai/glm-5.2 |
|---|---|---|---|
| Specs | |||
| Context window | 1M | 1M | 1M |
| Max output | 128K | 128K | 128K |
| Input modalities | TextImageVideo | Text | Text |
| Output modalities | Text | Text | Text |
| Capabilities | |||
| Streaming | Supported | Supported | Supported |
| Tool calling | Supported | Supported | Supported |
| Structured output (JSON) | Supported | Supported | Supported |
| Image input | Supported | Not supported | Not supported |
| Thinking mode | Supported | Supported | Supported |
| Price / 1M tokens | |||
| Input | $ 0.15$ 0.075 | $ 1.20 | $ 1.20 |
| Output | $ 0.50$ 0.25 | $ 4.40 | $ 4.40 |
| Cache read | $ 0.030$ 0.015 | $ 0.30 | $ 0.30 |
| Cache write | $ 0.15$ 0.075 | $ 1.20 | $ 1.20 |
| What it is good at | |||
| Best for |
|
|
|
| Published benchmarks | |||
| Terminal-Bench 2.1pass_ratesetups differ by: harness, max_output_tokens, temperature, top_p |
| — |
|
| Terminal-Bench 3.0pass_rate | — |
| — |
| Agents' Last Exam (CLI)scoresetups differ by: harness, max_output_tokens |
|
| — |
| AutomationBench v1.0.6score |
| — | — |
| DeepSWE 1.1resolved_rate |
|
| — |
| NL2Reposcore |
| — | — |
| SWE-Bench Proresolved_rate | — | — |
|
| CyberGymscore | — |
| — |
| ExploitBenchscore | — |
| — |
| ExploitGym (6h)tasks_completed | — |
| — |
| OfficeQA Proscore |
| — | — |
| GDPval-AA v2elo |
|
| — |
| AA Intelligence Index v4.1.1score |
| — | — |
| GPQA Diamondaccuracy | — | — |
|
| Humanity's Last Examaccuracy | — | — |
|
| Humanity's Last Exam (with tools)accuracysetups differ by: max_output_tokens, reasoning, temperature, tools, top_p |
|
|
|
| MCP Atlas Publicscore | — | — |
|
| Tool Decathlonscore | — | — |
|
| Toolathlon Verifiedpass_rate |
| — | — |
| MMVUaccuracy |
| — | — |