Qwen3 235B A22B Thinking 2507
Qwen · Open weights · Released Jul 25, 2025 · qwen/qwen3-235b-a22b-thinking-2507
Best $0.23 in / $2.30 out via Alibaba
131K contextToolsReasoningStructured outputOpen weights
Available from 3 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 6 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.23 | $2.30 | — | 131K | — | 1.7s | 80 tok/s | 100.0% | |
| $0.30 | $3.00 | — | 131K | fp8 | 3.3s | 35 tok/s | 100.0% | |
| $0.45 | $3.50 | — | 128K | fp8 | ✕ | ✕ | 99.3% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 6 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Recent changesRSS ↗
Aug 24, 2026context changed262144 → 131072
Aug 24, 2026endpoint removed4 → 3
About
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Specifications
Context window131,072
Max output117,964
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedJul 25, 2025
Knowledge cutoff2025-06-30
TokenizerQwen3
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-235b-a22b-thinking-2507)Sources
OpenRouter API23 minutes ago
AGGREGATORHugging Face23 minutes ago
COMMUNITYModelIndex benchmarks17 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.