Models/Qwen/Qwen3.8 2.4T A95B

Qwen3.8 2.4T A95B

Qwen · Open weights · Released Aug 12, 2026 · qwen/qwen3.8-2.4t-a95b
Best $2.00 in / $6.00 out via Alibaba · 1M ctx at this price (headline 1.05M)
CompareEstimate cost
1.05M contextToolsReasoningStructured outputPrompt cachingOpen weights
Available from 7 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 3 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSSMOKEUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · SMOKE: of 5 programmatic capability probes passed (JSON schema, instruction following, tool call, long-context, benign compliance; retry-once) · screened via OpenRouter routing · v0.2 · methodology v1.1 · measured 3 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$2.00 / $6.00
per MTok in / out
CACHE READ $0.25CACHE WRITE $2.50
Recent changesRSS ↗
Aug 29, 2026endpoint added · novita6 → 7
About

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...

Specifications
Context window1,048,576
Max output131,072
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedAug 12, 2026
TokenizerQwen
Open weights
Downloads33K
Likes1.3K
Licenseother
safetensors
Embed badge
Qwen3.8 2.4T A95B price and speed badge
[![Qwen3.8 2.4T A95B](https://modelindex.ai/badge/qwen/qwen3.8-2.4t-a95b)](https://modelindex.ai/models/qwen/qwen3.8-2.4t-a95b)
Sources
OpenRouter API23 minutes ago
AGGREGATOR
Hugging Face23 minutes ago
COMMUNITY
ModelIndex benchmarks20 days ago
MEASURED
Verified 23 minutes ago · list prices, not negotiated rates.