Models/DeepSeek V4.1 Flash vs GLM 5.2

DeepSeek V4.1 Flash vs GLM 5.2

API pricing, context, and independently measured performance, side by side. DeepSeek V4.1 Flash starts 82% cheaper on input than GLM 5.2 at list prices.

DeepSeek V4.1 Flash
DeepSeek · released Sep 10, 2026
Best input$0.10 / MTok
Best output$0.42 / MTok
Context1.05M
Providers28
Fastest measured0.64s° TTFT · 281 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
GLM 5.2
Z.ai · released Jun 16, 2026
Best input$0.561 / MTok
Best output$1.76 / MTok
Context1.05M
Providers31
Fastest measured0.61s° TTFT · 593 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
DeepSeek V4.1 FlashOPOpenInference$0.10$0.505s°10° tok/s
DeepSeek V4.1 FlashRelace logoRelace$0.10$0.50
DeepSeek V4.1 FlashDEDekaLLM$0.10$1.00
GLM 5.2Baidu logoBaidu$0.561$1.760.77s°90° tok/s3/5
GLM 5.2DeepInfra logoDeepInfra$0.563$1.801.2s°71° tok/s
GLM 5.2StreamLake logoStreamLake$0.641$2.011s°55° tok/s2/5
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 28 providers for DeepSeek V4.1 FlashAll 31 providers for GLM 5.2Open in comparator