模型精选78°

Deepseek 发布 V4.1-Flash 模型,参数量 5520 亿,内存需求减至前代四分之一

New Deepseek model V4.1-Flash cuts memory needs for AI agents

精选理由

Deepseek 新出的 V4.1-Flash 模型,参数量 5520 亿,内存占用比之前少四分之三,还比 GPT-5.6 Sol 和 Opus 5 表现好一点,适合做更便宜的 AI 代理。

Deepseek 推出 V4.1-Flash 多模态模型,参数量达 5520 亿,将 KV 缓存内存需求降至前代模型的四分之一。在 DeepSWE 编程基准测试中,该模型仅以 16 亿参数的活跃量,略胜于 Opus 5 和 GPT-5.6 Sol。该模型采用 MIT 许可证,旨在降低 AI 代理的成本。

原文 · Decoder

New Deepseek model V4.1-Flash cuts memory needs for AI agents

Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents. The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder .