DeepSeek V4 Flash 0731 vs GLM 5.3
API pricing, context, and independently measured performance, side by side. DeepSeek V4 Flash 0731 starts 96% cheaper on input than GLM 5.3 at list prices.
DeepSeek V4 Flash 0731
DeepSeek · released Jul 31, 2026
Best input$0.05 / MTok
Best output$0.16 / MTok
Context1.31M
Providers31
Fastest measured0.48s° TTFT · 65 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
GLM 5.3
Z.ai · released Aug 18, 2026
Best input$1.17 / MTok
Best output$3.96 / MTok
Context1.31M
Providers24
Fastest measured0.86s° TTFT · 52 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
| MODEL | PROVIDER | INPUT | OUTPUT | TTFT | TPS | SMOKE |
|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0731 | OPOpenInference | $0.05 | $0.16 | 14.9s° | 13° tok/s | 3/5 |
| DeepSeek V4 Flash 0731 | SRSail Research | $0.065 | $0.18 | 2s° | 41° tok/s | 5/5 |
| DeepSeek V4 Flash 0731 | AKAkashML | $0.065 | $0.18 | 1.1s° | 48° tok/s | — |
| GLM 5.3 | AKAkashML | $1.17 | $3.96 | — | — | — |
| GLM 5.3 | INIo Net | $1.19 | $4.18 | — | — | — |
| GLM 5.3 | $1.20 | $4.00 | — | — | — |
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page