Models/DeepSeek V4 Flash 0423 vs Gemma 4 31B

DeepSeek V4 Flash 0423 vs Gemma 4 31B

API pricing, context, and independently measured performance, side by side. DeepSeek V4 Flash 0423 starts 32% cheaper on input than Gemma 4 31B at list prices.

DeepSeek V4 Flash 0423
DeepSeek · released Apr 24, 2026
Best input$0.055 / MTok
Best output$0.109 / MTok
Context1.05M
Providers18
Fastest measured0.73s° TTFT · 138 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Gemma 4 31B
Google · released Apr 2, 2026
Best input$0.08 / MTok
Best output$0.34 / MTok
Context262K
Providers18
Fastest measured0.31s° TTFT · 205 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
DeepSeek V4 Flash 0423BABaidu$0.055$0.1090.73s°138° tok/s5/5
DeepSeek V4 Flash 0423STStreamLake$0.056$0.1122.2s°47° tok/s5/5
DeepSeek V4 Flash 0423GMIGMICloud$0.059$0.1182.5s°105° tok/s5/5
Gemma 4 31BOPOpenInference$0.08$0.352.3s°21° tok/s3/5
Gemma 4 31BDeepInfra logoDeepInfra$0.09$0.340.84s°56° tok/s5/5
Gemma 4 31BCWCoreWeave$0.10$0.341.6s°38° tok/s5/5
TTFT/TPS: ModelIndex-measured medians (P50, trailing 72h) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 18 providers for DeepSeek V4 Flash 0423All 18 providers for Gemma 4 31BOpen in comparator