DeepSeek V4 Flash 0731 vs GLM 5.3 Flash
API pricing, context, and independently measured performance, side by side. DeepSeek V4 Flash 0731 starts 60% cheaper on input than GLM 5.3 Flash at list prices.
DeepSeek V4 Flash 0731
DeepSeek · released Jul 31, 2026
Best input$0.03 / MTok
Best output$0.09 / MTok
Context1.31M
Providers30
Fastest measured—
Open weightsToolsReasoningStructured outputPrompt caching
GLM 5.3 Flash
Z.ai · released Aug 26, 2026
Best input$0.075 / MTok
Best output$0.25 / MTok
Context1.31M
Providers20
Fastest measured—
Open weightsVisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
| MODEL | PROVIDER | INPUT | OUTPUT | TTFT | TPS | SMOKE |
|---|---|---|---|---|---|---|
| DeepSeek V4 Flash 0731 | OPOpenInference | $0.03 | $0.10 | — | — | 3/5 |
| DeepSeek V4 Flash 0731 | BABaidu | $0.045 | $0.09 | — | — | 4/5 |
| DeepSeek V4 Flash 0731 | RERelace | $0.045 | $0.09 | — | — | 5/5 |
| GLM 5.3 Flash | RERelace | $0.075 | $0.25 | — | — | — |
| GLM 5.3 Flash | $0.075 | $0.25 | — | — | — | |
| GLM 5.3 Flash | $0.075 | $0.25 | — | — | — |