GLM 5.3 vs GLM 5.3 FlashX
API pricing, context, and independently measured performance, side by side. GLM 5.3 FlashX starts 59% cheaper on input than GLM 5.3 at list prices.
GLM 5.3
Z.ai · released Aug 18, 2026
Best input$0.892 / MTok
Best output$2.80 / MTok
Context1.31M
Providers34
Fastest measured0.61s° TTFT · 222 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
GLM 5.3 FlashX
Z.ai · released Sep 18, 2026
Best input$0.37 / MTok
Best output$1.25 / MTok
Context1.05M
Providers1
Fastest measured—
VisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
| MODEL | PROVIDER | INPUT | OUTPUT | TTFT | TPS | SMOKE |
|---|---|---|---|---|---|---|
| GLM 5.3 | BABaidu | $0.892 | $2.80 | — | — | — |
| GLM 5.3 | MOMorph | $0.893 | $2.81 | 3.9s° | 49° tok/s | — |
| GLM 5.3 | INInferenceNet | $0.90 | $3.00 | — | — | — |
| GLM 5.3 FlashX | $0.37 | $1.25 | — | — | — |
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page