Models/DeepSeek V4 Flash 0731 vs Kimi K3

DeepSeek V4 Flash 0731 vs Kimi K3

API pricing, context, and independently measured performance, side by side. DeepSeek V4 Flash 0731 starts 98% cheaper on input than Kimi K3 at list prices.

DeepSeek V4 Flash 0731
DeepSeek · released Jul 31, 2026
Best input$0.015 / MTok
Best output$0.132 / MTok
Context1.05M
Providers30
Fastest measured0.38s° TTFT · 265 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Kimi K3
Moonshot AI · released Jul 16, 2026
Best input$0.69 / MTok
Best output$11.25 / MTok
Context1.05M
Providers25
Fastest measured0.65s° TTFT · 100 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
DeepSeek V4 Flash 0731Relace logoRelace$0.015$1.280.69s°102° tok/s5/5
DeepSeek V4 Flash 0731OPOpenInference$0.015$1.281.9s°33° tok/s3/5
DeepSeek V4 Flash 0731SRSail Research$0.019$0.301.6s°50° tok/s5/5
Kimi K3Relace logoRelace$0.69$13.001.1s°150° tok/s—
Kimi K3Wafer logoWafer$0.70$16.000.87s°69° tok/s0/5
Kimi K3InferenceNet logoInferenceNet$0.72$13.0057.9s°7° tok/s—
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 30 providers for DeepSeek V4 Flash 0731All 25 providers for Kimi K3Open in comparator