Back to model catalog
Compare models
Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.
| Attribute | Hy4 Previewtencent/hy4-preview | Hy3tencent/hy3 |
|---|---|---|
| Specs | ||
| Context window | 1M | 262K |
| Max output | 66K | 131K |
| Input modalities | Text | Text |
| Output modalities | Text | Text |
| Capabilities | ||
| Streaming | Supported | Supported |
| Tool calling | Supported | Supported |
| Structured output (JSON) | Supported | Supported |
| Image input | Not supported | Not supported |
| Thinking mode | Supported | Supported |
| Price / 1M tokens | ||
| Input | $ 0.96 | $ 0.16 |
| Output | $ 2.88 | $ 0.64 |
| Cache read | $ 0.048 | $ 0.040 |
| Cache write | $ 0.96 | $ 0.16 |
| What it is good at | ||
| Best for |
|
|
| Published benchmarks | ||
| Terminal-Bench 2.1pass_ratesetups differ by: harness |
|
|
| BrowseComp (with tools)accuracy | — |
|
| Claw-Eval Generalpass3_rate | — |
|
| Agents' Last Exam (CLI)score |
| — |
| SWE-Bench Multilingualresolved_rate |
|
|
| SWE-Bench Proresolved_rate |
|
|
| SWE-Bench Verifiedresolved_rate | — |
|
| OfficeQA Proscore |
| — |
| GDPval-AA v2elo |
| — |
| AA-LCRscore | — |
|
| OneMillionBench (with tools)score |
| — |
| Humanity's Last Exam (with tools)accuracysetups differ by: tools |
|
|
| GPQA Diamondaccuracy |
|
|
| MCP Atlas Publicscore |
|
|
| Toolathlon Verifiedpass_rate |
| — |