Qwen3.5-27B
Qwen · Open weights · Released Feb 25, 2026 · qwen/qwen3.5-27b
262K contextVisionToolsReasoningStructured outputPrompt cachingOpen weights
Available from 6 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | SMOKE | UPTIME |
|---|---|---|---|---|---|---|---|---|---|
| $0.195 | $1.56 | — | 262K | — | 5s | 105 tok/s | 5/5 | 97.4% | |
| $0.25 | $2.00 | — | 262K | fp8 | 3.6s | 42 tok/s | 3/5 | 98.4% | |
| $0.26 | $2.60 | — | 262K | fp8 | 0.8s | 81 tok/s | 3/5 | 99.0% | |
| ATAtlasCloud | $0.27 | $2.16 | $0.27 | 262K | fp8 | 5.4s | 95 tok/s | 5/5 | 99.5% |
| PHPhala | $0.30 | $2.40 | $0.15 | 262K | — | 69.2s | 32 tok/s | 1/5 | 82.6% |
| $0.30 | $2.40 | — | 262K | bf16 | 5.2s | 88 tok/s | 4/5 | 99.9% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · SMOKE: of 5 programmatic capability probes passed (JSON schema, instruction following, tool call, long-context, benign compliance; retry-once) · screened via OpenRouter routing · v0.2 · methodology v1.1 · measured 43 hours ago
About
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
Specifications
Context window262,144
Max output65,536
Modalities intext, image, video
Modalities outtext
LicenseOpen weights
ReleasedFeb 25, 2026
TokenizerQwen3
WeightsHugging Face ↗
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3.5-27b)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks43 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.