Qwen3.8 Flash
Qwen · Open weights · Released Aug 26, 2026 · qwen/qwen3.8-flash
Best $0.15 in / $0.47 out via Alibaba
1M contextVisionToolsReasoningStructured outputPrompt cachingOpen weights
Available from 2 providersUSD / MTOK · PRICES 27 MINUTES AGO · SPEED 4 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.15 | $0.47 | $0.016 | 1M | — | ✕ | ✕ | 100.0% | |
| $0.15 | $0.47 | $0.016 | 1M | — | — | — | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 4 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Input price history · best listed price, observed
↓ 6% SINCE AUG 29, 2026
$0.16 · AUG 29, 2026$0.15 · TODAY
Pricing variants · USD / MTok in / out
STANDARD
$0.15 / $0.47
per MTok in / out
CACHE READ $0.016
Recent changesRSS ↗
Sep 16, 2026endpoint removed · makora2 → 1
Sep 10, 2026endpoint added · makora1 → 2
Aug 29, 2026price decreased · input0.16 → 0.15
Aug 27, 2026model added
About
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
Specifications
Context window1,000,000
Max output131,072
Modalities intext, image, video
Modalities outtext
LicenseOpen weights
ReleasedAug 26, 2026
TokenizerQwen
WeightsHugging Face ↗
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3.8-flash)Sources
OpenRouter API27 minutes ago
AGGREGATORHugging Face27 minutes ago
COMMUNITYModelIndex benchmarks4 days ago
MEASUREDVerified 27 minutes ago · list prices, not negotiated rates.