DeepSeek announced Monday that its upcoming V4.1 Flash model has outperformed the V4 Pro across all key benchmarks during extensive internal and external testing. The company said it will officially launch the model around September 10, 2026 (Beijing Time).
In a notable shift, DeepSeek will redirect all requests to the Pro model to V4.1 Flash following the launch, billing them at Flash rates until the V4.1 Pro is released. The company invited users to report any issues encountered during comparative testing between the two models.
DeepSeek also detailed new pricing for the Flash series effective September 10. During off-peak hours, rates are $0.003 for input cache hits, $0.15 for input cache misses, and $0.6 for output. Peak-hour prices will be double those rates. The move signals a continued price war in the AI model market, where DeepSeek has positioned itself as a lower-cost alternative to Western competitors.