โŒ

Normal view

New Deepseek model V4.1-Flash cuts memory needs for AI agents

10 September 2026 at 12:40

Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents.

The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder.

Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks

21 August 2026 at 19:08

Deepseek has released V4-Flash-Vision-Exp, an experimental multimodal model that adds image understanding to V4-Flash's text capabilities. On the company's own multimodal agent benchmarks, it approaches Opus 4.8 and sometimes beats it.

The article Deepseek releases experimental Flash vision model that rivals Opus 4.8 on agent benchmarks appeared first on The Decoder.

โŒ