Models/DeepSeek V4.1 Flash vs GLM 5.3 Flash

DeepSeek V4.1 Flash vs GLM 5.3 Flash

API pricing, context, and independently measured performance, side by side.

DeepSeek V4.1 Flash
DeepSeek · released Sep 10, 2026
Best input$0.02 / MTok
Best output$0.39 / MTok
Context1.05M
Providers33
Fastest measured0.29s° TTFT · 210 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
GLM 5.3 Flash
Z.ai · released Aug 26, 2026
Best input$0.02 / MTok
Best output$0.10 / MTok
Context1.31M
Providers37
Fastest measured0.49s° TTFT · 184 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
DeepSeek V4.1 FlashRelace logoRelace$0.02$0.601.2s°23° tok/s—
DeepSeek V4.1 FlashOPOpenInference$0.03$0.505s°10° tok/s—
DeepSeek V4.1 FlashInferenceNet logoInferenceNet$0.06$0.39———
GLM 5.3 FlashOPOpenInference$0.02$0.302.6s°47° tok/s—
GLM 5.3 FlashRelace logoRelace$0.025$0.500.93s°34° tok/s—
GLM 5.3 FlashPareto Inference logoPareto Inference$0.03$0.102.6s70 tok/s—
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 33 providers for DeepSeek V4.1 FlashAll 37 providers for GLM 5.3 FlashOpen in comparator