❌

Normal view

New Deepseek model V4.1-Flash cuts memory needs for AI agents

10 September 2026 at 12:40

Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents.

The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder.

Deepseek plans the largest known Huawei chip cluster with 160,000 processors in Inner Mongolia

4 September 2026 at 14:19

Deepseek wants to put 160,000 Huawei Ascend-950DT chips into an Inner Mongolia data center for inference only, not training. It would be the largest known Huawei chip cluster. But production bottlenecks mean Huawei probably can't deliver for over a year.

The article Deepseek plans the largest known Huawei chip cluster with 160,000 processors in Inner Mongolia appeared first on The Decoder.

❌