Qwen3 Next 80B A3B Instruct
Qwen · Open weights · Released Sep 11, 2025 · qwen/qwen3-next-80b-a3b-instruct
Best $0.09 in / $0.78 out via DeepInfra · fp8
262K contextToolsStructured outputPrompt cachingOpen weights
Available from 4 providersUSD / MTOK · PRICES 25 MINUTES AGO · SPEED 24 HOURS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.09 | $1.10 | — | 262K | fp8 | 0.5s | 31 tok/s | 99.6% | |
| $0.098 | $0.78 | — | 131K | — | ✕ | ✕ | 100.0% | |
| $0.10 | $1.10 | $0.07 | 262K | fp8 | ✕ | ✕ | 100.0% | |
| $0.15 | $1.20 | — | 262K | — | 0.63s | 348 tok/s | 99.9% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 24 hours ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Recent changesRSS ↗
Oct 9, 2026endpoint removed · novita5 → 4
About
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
Specifications
Context window262,144
Max output16,384
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedSep 11, 2025
Knowledge cutoff2025-09-30
TokenizerQwen3
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-next-80b-a3b-instruct)Sources
OpenRouter API25 minutes ago
AGGREGATORHugging Face25 minutes ago
COMMUNITYModelIndex benchmarks20 days ago
MEASUREDVerified 25 minutes ago · list prices, not negotiated rates.