deepseek-v4-flash-0731
Alibaba Cloud Int
deepseek-v4-flash-0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows. This is the GA release of DeepSeek V4 Flash.
Input / output per 1M
$0.424 / $1.272
Context Window
1M
Max output
384,000
Released
Jul 31, 2026
Provider pricing
The same model is offered by several providers. Prices below are per provider; select a row to switch provider.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output |
|---|---|---|---|---|---|
Current | $0.424 | $1.272 | $0.042 | 1M | 384,000 |
Channel price | $0.44 | $1.32 | $0.014 | 1M | 384,000 |
Pricing
Prices from the current provider, per 1M tokens.
| Input / 1M | Output / 1M | Cache read / 1M |
|---|---|---|
| $0.424 | $1.272 | $0.042 |
Performance
Data for this section is coming soon.
Uptime
Data for this section is coming soon.