Skip to main content
ModelsEnterpriseResearch

DeepSeek releases V4.1-Flash, says it outperforms flagship V4-Pro

Source: SiliconANGLE (opens in a new tab) · Duncan Riley

Intel Summary

Hangzhou DeepSeek Artificial Intelligence Basic Technology Research Co. Ltd. has released DeepSeek-V4.1-Flash, an open-weight model that is the smallest variant in a new architecture family. The company claims third-party evaluations demonstrate that V4.1-Flash outperforms its larger DeepSeek-V4-Pro model in performance, speed, total runtime, and cost.

Why It Matters

If validated, a smaller open-weight architecture exceeding larger flagship performance offers significant inference cost reductions and latency improvements. Teams building on DeepSeek infrastructure must evaluate V4.1-Flash ahead of stated operational changes affecting V4-Pro requests starting September 14.

Part of an ongoing development

Source

DeepSeek released V4.1-Flash

DeepSeek has released V4.1-Flash, a multimodal model containing 552 billion total parameters with 16 billion active parameters per token. According to benchmark results reported from DeepSeek, V4.1-Flash narrowly outperforms Opus 5 and GPT-5.6 Sol on the DeepSWE coding benchmark and is published under an open MIT license. Claims are as reported; this summary makes no determination about accuracy or significance.

More coverage of this development

Organizations & Entities

Topics