Models/DeepSeek/DeepSeek V4 Flash 0423

DeepSeek V4 Flash 0423

✓ VERIFIED BY DEEPSEEK
DeepSeek · Open weights · Released Apr 24, 2026 · deepseek/deepseek-v4-flash
Best $0.017 in / $0.14 out via Relace · fp4
CompareEstimate cost
1.05M contextToolsReasoningStructured outputPrompt cachingOpen weights
Available from 20 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 9 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSSMOKEUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured against the provider API directly; ° = routed via OpenRouter (adds a hop) · ✕ = probed, endpoint returned no measurable stream · SMOKE: of 5 programmatic capability probes passed (JSON schema, instruction following, tool call, long-context, benign compliance; retry-once) · screened via OpenRouter routing · v0.2 · methodology v1.1 · measured 9 minutes ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Quality benchmarksCURATED AUG 24, 2026
LMArena Elo1436INDEPENDENTEntry: deepseek-v4-flash (rank 86)source ↗ 2026-08-21
Third-party reported results, hand-curated with sources — not measured by ModelIndex. Our own measured serving data (TTFT/TPS/SMOKE) is in the providers table above.
Input price history · best listed price, observed
↓ 76% SINCE AUG 22, 2026
$0.068 · AUG 22, 2026$0.017 · TODAY
Pricing variants · USD / MTok in / out
STANDARD
$0.017 / $1.28
per MTok in / out
CACHE READ $0.017CACHE WRITE Free
Recent changesRSS ↗
Oct 8, 2026endpoint added · wafer19 → 20
Oct 6, 2026price decreased · input0.03 → 0.0082
Oct 5, 2026price increased · input0.0224 → 0.03
Oct 5, 2026price increased · output0.056 → 0.084
Oct 4, 2026endpoint added · cloudflare18 → 19
Oct 3, 2026price increased · input0.0003 → 0.0081
Oct 3, 2026price decreased · output0.0837 → 0.056
Oct 2, 2026price increased · input0.0002 → 0.0003
About

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Specifications
Context window1,048,576
Max output943,718
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedApr 24, 2026
TokenizerDeepSeek
Open weights
Downloads995.9K
Likes2.3K
Licensemit
safetensors
Embed badge
DeepSeek V4 Flash 0423 price and speed badge
[![DeepSeek V4 Flash 0423](https://modelindex.ai/badge/deepseek/deepseek-v4-flash)](https://modelindex.ai/models/deepseek/deepseek-v4-flash)
Sources
deepseek ↗23 minutes ago
OFFICIAL PRICING
OpenRouter API23 minutes ago
AGGREGATOR
LiteLLM dataset23 minutes ago
AGGREGATOR
Hugging Face23 minutes ago
COMMUNITY
ModelIndex benchmarks2 days ago
MEASURED
✓ Verified against DeepSeek’s own source.
Verified 23 minutes ago · list prices, not negotiated rates.