Qwen · Model guide
Qwen3.6 Flash
A fast 1M-context native multimodal Qwen model for agentic coding, visual understanding and high-volume tool-driven applications, with optional reasoning and structured output.
What this model is good at
What the vendor positions it for, and which of our models to reach for instead.
VendorA speed-oriented multimodal Agent model covering coding, reasoning and visual tasks.source
Fast multimodal agents · Designed for low-latency agent coding and visual workflows at large context sizes.(vendor claim)source
- qwen/qwen3.7-flashsuccessor —Choose Qwen3.7 Flash for the current Flash generation.(vendor claim)source
Pricing and billing
Tiered pricing: the rate changes once the input passes the threshold.
| Price / 1M tokens by input size | Input < 256.0K | Input ≥ 256.0K |
|---|---|---|
| Input | $ 0.19 | $ 0.75 |
| Output | $ 1.13 | $ 4.50 |
| Cache read | $ 0.019 | $ 0.075 |
| Cache write | $ 0.24 | $ 0.94 |
At this input size you are on the Input < 256.0K tier.
+ 2.0K × $ 1.13
Capabilities and limits
What CrossModel guarantees across every route this model can take right now.
- Context window
- 1.0M tokens
- Max output
- 65.5K tokens
- Input / output modalities
- Text + Image + Video → Text
- Streaming
- Supported
- Tool calling
- Supported
- Structured output (JSON)
- Supported
- Image input
- Supported
- Thinking mode
- Supported· can be turned off
- Available endpoints
- /v1/chat/completions · /v1/responses · /v1/messages
Published benchmarks
Scores the vendor reported, with the evaluation setup each one came from.
- SWE-Bench Proresolved_rate·Qwen Team · 2026-04-15setup not fully disclosed
Refined public task set; internal agent scaffold, 200K context.
49.5% - SWE-Bench Verifiedresolved_rate·Qwen Team · 2026-04-15setup not fully disclosed
Qwen states that Qwen3.6-35B-A3B is served through the API as Qwen3.6-Flash; internal agent scaffold, 200K context.
73.4%
Vendor-reported numbers, not CrossModel measurements. Scores are only comparable when the benchmark version, metric and evaluation setup match, so nothing here is averaged or ranked.
Use it in your tools
Point the base URL at CrossModel and paste this model ID — every tool below has a setup guide.
Frequently asked questions
What is Qwen3.6 Flash?
How much does Qwen3.6 Flash cost?
Does Qwen3.6 Flash support tool calling and structured output?
Which endpoint do I call?
Can I try it without writing code?
Chat first, integrate later
Chat first, then wire it in once you like the answers.