模型官方一手精选78°

DeepSeek 发布 V4.1-Flash 模型,支持百万级上下文

DeepSeek-V4.1-Flash Ships Causal Encoder–Decoder MoE With 1M Context and Extreme KV Compression

精选理由

DeepSeek 新出的 V4.1-Flash 模型,上下文能到百万级,比之前版本长很多,适合处理长文本对话。

DeepSeek 的 V4.1-Flash 模型采用 552B 的因果编码器-解码器 MoE 结构,上下文长度可达 1M,通过 CSA2+FP4 技术实现每 token 约 890 字节的 KV 缓存压缩。

原文 · pandaily

DeepSeek-V4.1-Flash Ships Causal Encoder–Decoder MoE With 1M Context and Extreme KV Compression

DeepSeek’s V4.1-Flash pairs a 552B Causal Encoder–Decoder MoE with 1M context, CSA2+FP4 KV at ~890 bytes/token, and API access as deepseek-flash.