Models/Qwen/Qwen3.8 Flash

Qwen3.8 Flash

Qwen · Open weights · Released Aug 26, 2026 · qwen/qwen3.8-flash
Best $0.15 in / $0.47 out via Alibaba
CompareEstimate cost
1M contextVisionToolsReasoningStructured outputPrompt cachingOpen weights
Available from 2 providersUSD / MTOK · PRICES 27 MINUTES AGO · SPEED 4 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 4 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Input price history · best listed price, observed
↓ 6% SINCE AUG 29, 2026
$0.16 · AUG 29, 2026$0.15 · TODAY
Pricing variants · USD / MTok in / out
STANDARD
$0.15 / $0.47
per MTok in / out
CACHE READ $0.016
Recent changesRSS ↗
Sep 16, 2026endpoint removed · makora2 → 1
Sep 10, 2026endpoint added · makora1 → 2
Aug 29, 2026price decreased · input0.16 → 0.15
Aug 27, 2026model added
About

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Specifications
Context window1,000,000
Max output131,072
Modalities intext, image, video
Modalities outtext
LicenseOpen weights
ReleasedAug 26, 2026
TokenizerQwen
Open weights
Downloads1.8M
Likes6.1K
Licenseother
safetensors
Embed badge
Qwen3.8 Flash price and speed badge
[![Qwen3.8 Flash](https://modelindex.ai/badge/qwen/qwen3.8-flash)](https://modelindex.ai/models/qwen/qwen3.8-flash)
Sources
OpenRouter API27 minutes ago
AGGREGATOR
Hugging Face27 minutes ago
COMMUNITY
ModelIndex benchmarks4 days ago
MEASURED
Verified 27 minutes ago · list prices, not negotiated rates.