Qwen3 30B A3B Instruct 2507
Qwen · Open weights · Released Jul 29, 2025 · qwen/qwen3-30b-a3b-instruct-2507
262K contextToolsStructured outputPrompt cachingOpen weights
Available from 5 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| STStreamLake | $0.048 | $0.193 | — | 128K | — | 1.2s | 215 tok/s | 99.6% |
| $0.09 | $0.30 | — | 262K | fp8 | 2.4s | 27 tok/s | 98.8% | |
| CWCoreWeave | $0.10 | $0.30 | $0.10 | 262K | bf16 | 0.38s | 172 tok/s | 99.8% |
| $0.10 | $0.30 | — | 262K | fp8 | 0.51s | 139 tok/s | 97.5% | |
| $0.13 | $0.52 | — | 131K | — | 0.82s | 106 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 43 hours ago
About
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and...
Specifications
Context window262,144
Max output32,000
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedJul 29, 2025
Knowledge cutoff2025-06-30
TokenizerQwen3
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-30b-a3b-instruct-2507)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks43 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.