DeepSeek on Sept. 10 formally released its V4.1 Flash model and introduced new Flash-series pricing, according to a notification sent to API users.

DeepSeek said that V4.1 Flash had surpassed V4 Pro in performance, cost, speed and total completion time in internal and external testing. Starting Sept. 14 at noon Beijing time, requests to V4 Pro will be routed to V4.1 Flash and billed at Flash-series rates.

During off-peak hours, the new pricing is RMB 0.02 per million tokens for cache-hit input, RMB 1 for cache-miss input and RMB 4 for output. Peak-hour prices will be twice as high. [IT Home, in Chinese]