On September 8, 2026, DeepSeek announced the start of beta testing for the V4.1 Flash intermediate version. This version features a new model structure, supports native multimodality, and offers faster processing times and lower costs. It performed nearly on par with Opus-4.8 in the multimodal AgentBenchmark. Test feedback showed that its end-to-end speed was significantly better than the previous version: SVG code generation was 6 times faster, and long context retrieval was more than 5 times faster. The beta testing period will end automatically on September 10.
However, on September 9, DeepSeek announced that it would adjust the pricing of the Flash series models starting the next day (September 10), with a maximum discount of 60%. After this adjustment, the unit price for cache hits during idle periods dropped to 0.02 yuan, for missed hits to 1 yuan, and for output operations to 4 yuan; prices during peak hours doubled accordingly. Compared to the previous pricing levels, comprehensive estimates indicate that developer costs may decrease by about 40%, and some prices have already…
DeepSeek’s new model is available for limited-time beta testing! It’s faster, and the cost of using it is also lower.
On September 8th, DeepSeek announced that the V4.1 Flash intermediate version was starting beta testing. This model features a new structure and supports native multi-modal processing, making it faster and less expensive. Developers need to call the model with the name deepseek-v4.1-flash-expires-on-0910.
3 reports
2026-09-09
DeepSeek announced a price reduction! Up to 60% off!
On September 9th, DeepSeek announced that starting at 12:00 on September 10, 2026, Beijing time, the pricing of the flash series models would be adjusted, with up to 60% reductions. After this adjustment, the unit price for cache hits during idle periods will be 0.02 yuan, and for misses, it will be 1 yuan. The output unit price will be…