header-langage
简体中文
繁體中文
English
Tiếng Việt
한국어
日本語
ภาษาไทย
Türkçe
Scan to Download the APP

The old Pro scored only 8 points, while the DeepSeek V4 Flash surged to 54.4 points in DeepSWE, approaching the Opus 4.8.

According to Dynamic Insight Beating monitoring, on the DeepSWE long-range software engineering benchmark, DeepSeek-V4-Flash Official Version scored 54.4 points, approaching Claude Opus 4.8 with 59 points, far exceeding the previous 8 points of V4-Pro-Preview.

This score is lower than Kimi K3's 69 points but higher than GLM-5.2's 44 points, placing it in the top tier of domestic models.

However, this is not a raw model test under exactly the same conditions. The official V4-Flash version uses DeepSeek's as-yet-unreleased Harness minimalist mode, where the Agent's performance is affected by both the model and execution framework.

举报 Correction/Report
Correction/Report
Submit
Add Library
Visible to myself only
Public
Save
Choose Library
Add Library
Cancel
Finish