Back to model catalog
Compare models
Context, limits, capabilities, price and use cases side by side — up to 4 models at a time. The URL holds your selection, so you can share it.
| Attribute | Kimi K2.6moonshot/kimi-k2.6 | Kimi K2.7 Codemoonshot/kimi-k2.7-code | Kimi K3moonshot/kimi-k3 |
|---|---|---|---|
| Specs | |||
| Context window | 262K | 262K | 1M |
| Max output | 262K | 262K | 1M |
| Input modalities | TextImageVideo | TextImageVideo | TextImageVideo |
| Output modalities | Text | Text | Text |
| Capabilities | |||
| Streaming | Supported | Supported | Supported |
| Tool calling | Supported | Supported | Supported |
| Structured output (JSON) | Supported | Supported | Supported |
| Image input | Supported | Supported | Supported |
| Thinking mode | Supported | Supported | Supported |
| Price / 1M tokens | |||
| Input | $ 1.00 | $ 1.00 | $ 3.00 |
| Output | $ 4.16 | $ 4.16 | $ 15.00 |
| Cache read | $ 0.18 | $ 0.18 | $ 0.30 |
| Cache write | $ 1.00 | $ 1.00 | $ 3.00 |
| What it is good at | |||
| Best for |
|
|
|
| Published benchmarks | |||
| Kimi Code Bench v2score | — |
| — |
| Terminal-Bench 2.0pass_rate |
| — | — |
| Terminal-Bench 2.1pass_rate | — | — |
|
| BrowseComp (with tools)accuracysetups differ by: reasoning |
| — |
|
| DeepSearchQAf1 |
| — |
|
| Kimi Claw 24/7score | — |
| — |
| DeepSWEresolved_rate | — | — |
|
| MLS Bench Litescore | — |
| — |
| ProgramBenchscoresetups differ by: harness | — |
|
|
| SWE-Bench Proresolved_rate |
| — | — |
| SWE-Bench Verifiedresolved_rate |
| — | — |
| OSWorld Verifiedsuccess_rate |
| — |
|
| MMMU Proaccuracy |
| — | — |
| Humanity's Last Exam Full (with tools)accuracysetups differ by: reasoning, tools |
| — |
|
| GPQA Diamondaccuracysetups differ by: reasoning |
| — |
|
| MCP Atlasscore | — |
| — |
| MCPMark Verifiedscoresetups differ by: max_output_tokens, runs | — |
|
|