Qwen3.5-Flash
Qwen · Proprietary · Released Feb 25, 2026 · qwen/qwen3.5-flash-02-23
1M contextVisionToolsReasoningStructured output
Available from 1 providerUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.065 | $0.26 | — | 1M | — | 1.1s | 151 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 43 hours ago
About
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the...
Specifications
Context window1,000,000
Max output65,536
Modalities intext, image, video
Modalities outtext
LicenseProprietary
ReleasedFeb 25, 2026
TokenizerQwen3
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3.5-flash-02-23)Sources
OpenRouter API4 minutes ago
AGGREGATORModelIndex benchmarks43 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.