B · Normal
[DeepSeek V4.1 Flash model officially released] On September 10th, DeepSeek officially released the DeepSeek V4.1 Flash model. It is the smallest model in the company’s new model structure series, featuring native multi-modal visual understanding. According to reports, the new generation model significantly reduces the size of the KV Cache. Compared with the previous generation model, the demand for HBM is reduced to 1/4 and the demand for SSD is reduced to 1/8. In Agent usage scenarios, the cost of cache hits often accounts for a relatively high proportion, and the compression of KV Cache greatly reduces the usage cost of Agent tasks.
Futures Market Intelligence 🕐 2026-09-10 14:07

Related telegraphs

Comments

0/500
Captcha (click to refresh)
No comments yet