DeepSeek is scheduled to release its V4.1 Flash model around September 10, 2026 (Beijing Time). This new model has undergone extensive internal and external testing.
The V4.1 Flash model has demonstrated superior performance across key metrics, including overall performance, cost efficiency, processing speed, and task completion time, compared to the existing V4 Pro model.
Following the official launch of V4.1 Flash and prior to the release of V4.1 Pro, all requests directed to the V4 Pro model will be automatically rerouted to V4.1 Flash. These requests will also be billed at the V4.1 Flash pricing.
This change aims to provide users with the benefits of the newer, more capable model at a reduced cost without requiring manual migration.
DeepSeek will implement new pricing for its Flash series, effective from 12:00 Beijing Time on September 10, 2026. During off-peak hours, the unit price for input cache hits will be $0.003, input cache misses $0.15, and output $0.6.
Peak-hour usage will incur charges that are double the off-peak rates. Users are advised to consider these new rates when planning their model usage.
✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →
One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.
One email a day. Unsubscribe in one click, any time.
Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.
▶ Play today's briefNew every morning, and the back catalogue is archived by date.
DeepSeek will release its V4.1 Flash AI model around September 10, 2026, which has outperformed the V4 Pro model in performance, cost, speed, and task completion. All V4 Pro requests will be routed to V4.1 Flash and billed at the lower Flash price after the launch, ahead of V4.1 Pro's release.