Qwen3 235B A22B Thinking 2507
Qwen · Open weights · Released Jul 25, 2025 · qwen/qwen3-235b-a22b-thinking-2507
131K contextToolsReasoningStructured outputOpen weights
Available from 3 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.23 | $2.30 | — | 131K | — | 1.7s | 81 tok/s | 100.0% | |
| $0.30 | $3.00 | — | 131K | fp8 | 101.9s | 17 tok/s | 99.5% | |
| $0.45 | $3.50 | — | 128K | fp8 | 1.4s | 40 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 43 hours ago
Recent changesRSS ↗
Aug 24, 2026context changed262144 → 131072
Aug 24, 2026endpoint removed4 → 3
About
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Specifications
Context window131,072
Max output—
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedJul 25, 2025
Knowledge cutoff2025-06-30
TokenizerQwen3
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-235b-a22b-thinking-2507)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks43 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.