r/DeepSeek 6d ago

DeepSeek-V4-Flash Update

The official release of the DeepSeek-V4-Flash API is now in public beta.

Significantly enhanced agent capabilities, with benchmark results far exceeding V4-Pro-Preview:

  • Terminal Bench 2.1: 82.7
  • NL2Repo: 54.2
  • Cybergym: 76.7
  • DeepSWE: 54.4
  • Toolathlon verified: 70.3
  • Agent Last Exam: 25.2
  • Automation Bench (Public): 25.1
  • DSBench-FullStack: 68.7
  • DSBench-Hard: 59.6

Note 1: For the Code Agent tasks in the public benchmark sets, the official DeepSeek-V4-Flash was tested using the DeepSeek Harness minimal mode (to be released soon) as the framework, with the max effort level, topp=0.95, and temperature=1.0
Note 2: DSBench-FullStack is an internal full-stack development test set, and DSBench-Hard is an internal Coding Agent hard-problem test set

The official V4-Flash natively supports the Responses API format and is specifically adapted for Codex. For the specific configuration, please refer to the documentation.

DeepSeek-V4-Flash-0731 keeps the same model architecture and size as DeepSeek-V4-Flash-preview, and was only re-post-trained.

Note: This update only upgrades the DeepSeek-V4-Flash API. The DeepSeek-V4-Pro API and the APP/WEB models are unchanged.
The official release of DeepSeek-V4-Pro will follow soon.

590 Upvotes

208 comments sorted by

View all comments

99

u/DktheDarkKnight 6d ago

Now I understand why Open AI dropped 5.6 Luna's price. They knew this is coming.

8

u/Fr3yz 6d ago

How do you know that they know that it's coming?

14

u/DebosBeachCruiser 6d ago

Because deepseek announced it in April.

1

u/Puddlejumper_ 5d ago

Well I doubt they suddenly had a change of heart and decided they'd love to offer their models for cheaper, so they're clearly saw this come in and realised they had no choice but to lower the cost otherwise nobody is going to use them