modelbenchmark.io

DeepSeek V4 Flash (Speed)

DeepSeek · deepseek-v4-flash-speed

Compare

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding

Specification

most-agreed values

Context
1M
Max output
393K
Released
2026-07-31
Knowledge cutoff
2025-05
Retires
—
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.14
Output
$0.28
Cache read
$0.028

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
neuralwatt$0.14$0.28$0.028—1M393K—

More from DeepSeek

most-hosted first