New model repo: deepseek-ai/DeepSeek-V4-Flash-0731
text generation · 4,559,659 downloads · 3,879 likes · tags: deepseek_v4, text-generation, conversational, eval-results, endpoints_compatible, 8-bit, fp8
Model Watch · DeepSeek
Forty-four cents a million on a million-token window, twenty-two off-peak. The open weights went up the same day.
DeepSeek-V4-Flash is a model from DeepSeek, built in China. It's out, and you can call it today.
DeepSeek released it on 31 July 2026. That date comes from the lab's own docs.
The tracker first logged the name on 4 July 2026 and last saw it on 1 August 2026.
| Lab | DeepSeek |
|---|---|
| Built in | China |
| Status | released |
| API model id | deepseek-v4-flash |
| Released | 31 July 2026 |
| Context window | 1M tokens |
| Max output | 384K tokens |
| Input, per 1M tokens | $0.44 |
| Output, per 1M tokens | $1.32 |
List prices per million tokens, read off the lab's own docs. Caching, batching and volume deals all cut the real number, so treat this as the ceiling. Check the source before you budget against it.
| Model | Status | Date | Context | In / 1M | Out / 1M |
|---|---|---|---|---|---|
| DeepSeek-V4-Flash | released | 31 July 2026 | 1M tokens | $0.44 | $1.32 |
| DeepSeek-V4-Flash-Vision-Exp | preview | 21 August 2026 | 1M tokens | $0.44 | $1.32 |
| DeepSeek-V4-Pro | released | 13 August 2026 | 1M tokens | $1.32 | $3.96 |
A dash means the lab hasn't published that number for that model.
We tracked 3 public discussions naming DeepSeek-V4-Flash across r/ClaudeCode, r/DeepSeek and r/accelerate between 9 August 2026 and 27 August 2026.
The subject that came up most was what it costs, then using it to write code. A discussion can cover several, so these add up to more than 3.
| What people are discussing | Threads |
|---|---|
| what it costs | 3 |
| using it to write code | 2 |
Every tracked post, release, changelog entry or incident whose title or summary names DeepSeek-V4-Flash. Newest first, each one linking to where the lab published it.
text generation · 4,559,659 downloads · 3,879 likes · tags: deepseek_v4, text-generation, conversational, eval-results, endpoints_compatible, 8-bit, fp8
text generation · 498,733 downloads · 268 likes · tags: deepseek_v4, text-generation, endpoints_compatible, 8-bit, fp8
The release timeline is every model we know about, filterable by lab, country and date. The tracker is the live feed these mentions come from, and Model Watch is the whole catalogue on one page.
More from DeepSeek:
.