DeepSeek 发布 V4.1-Flash 模型,支持百万级上下文
DeepSeek-V4.1-Flash Ships Causal Encoder–Decoder MoE With 1M Context and Extreme KV Compression
DeepSeek 新出的 V4.1-Flash 模型,上下文能到百万级,比之前版本长很多,适合处理长文本对话。
DeepSeek 的 V4.1-Flash 模型采用 552B 的因果编码器-解码器 MoE 结构,上下文长度可达 1M,通过 CSA2+FP4 技术实现每 token 约 890 字节的 KV 缓存压缩。
DeepSeek-V4.1-Flash Ships Causal Encoder–Decoder MoE With 1M Context and Extreme KV Compression
DeepSeek’s V4.1-Flash pairs a 552B Causal Encoder–Decoder MoE with 1M context, CSA2+FP4 KV at ~890 bytes/token, and API access as deepseek-flash.