Models/DeepSeek V4 Flash 0731 vs GLM 5.3

DeepSeek V4 Flash 0731 vs GLM 5.3

API pricing, context, and independently measured performance, side by side. DeepSeek V4 Flash 0731 starts 96% cheaper on input than GLM 5.3 at list prices.

DeepSeek V4 Flash 0731
DeepSeek · released Jul 31, 2026
Best input$0.05 / MTok
Best output$0.16 / MTok
Context1.31M
Providers31
Fastest measured0.48s° TTFT · 65 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
GLM 5.3
Z.ai · released Aug 18, 2026
Best input$1.17 / MTok
Best output$3.96 / MTok
Context1.31M
Providers24
Fastest measured0.86s° TTFT · 52 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
DeepSeek V4 Flash 0731OPOpenInference$0.05$0.1614.9s°13° tok/s3/5
DeepSeek V4 Flash 0731SRSail Research$0.065$0.182s°41° tok/s5/5
DeepSeek V4 Flash 0731AKAkashML$0.065$0.181.1s°48° tok/s
GLM 5.3AKAkashML$1.17$3.96
GLM 5.3INIo Net$1.19$4.18
GLM 5.3DeepInfra logoDeepInfra$1.20$4.00
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 31 providers for DeepSeek V4 Flash 0731All 24 providers for GLM 5.3Open in comparator