Qwen3.8 Max (0902) vs GLM 5.3
API pricing, context, and independently measured performance, side by side. GLM 5.3 starts 43% cheaper on input than Qwen3.8 Max (0902) at list prices.
Qwen3.8 Max (0902)
Qwen · released Sep 3, 2026
Best input$2.00 / MTok
Best output$6.00 / MTok
Context1M
Providers1
Fastest measured—
VisionToolsReasoningStructured outputPrompt caching
GLM 5.3
Z.ai · released Aug 18, 2026
Best input$1.15 / MTok
Best output$3.50 / MTok
Context1.31M
Providers26
Fastest measured0.86s° TTFT · 52 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
| MODEL | PROVIDER | INPUT | OUTPUT | TTFT | TPS | SMOKE |
|---|---|---|---|---|---|---|
| Qwen3.8 Max (0902) | $2.00 | $6.00 | — | — | — | |
| GLM 5.3 | REReka | $1.15 | $3.50 | 1.8s° | 15° tok/s | — |
| GLM 5.3 | AKAkashML | $1.17 | $3.96 | — | — | — |
| GLM 5.3 | DEDecart | $1.19 | $3.74 | — | — | — |
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page