DeepSeek 发布多模态模型 V4.1-Flash,KV cache 内存降至四分之一
New Deepseek model V4.1-Flash cuts memory needs for AI agents
阅读原文
本站未展示全文,请前往来源网站阅读。
AI 导读
DeepSeek 发布多模态模型 V4.1-Flash,总参数 552B,每个 token 仅激活 16B,KV cache 内存降至前代的四分之一。该模型以 MIT 许可证发布,在 DeepSWE 编码基准上略微超过 Opus 5 和 GPT-5.6 Sol,面向更低成本的 AI 智能体。
来源:The Decoder · the-decoder.com