SiliconFlow· @SiliconFlowAI · X·· 22 天前精选AI 评分67
AI 导读
DeepSeek-V4.1-Flash 在 SiliconFlow 实现 Day 0 上线。该模型为 552B MoE,prefill 约激活 8B、decode 约激活 16B,支持原生视觉和 1M 上下文窗口,KV cache 占用相比 V4 Flash 缩小约 4 倍,采用 MIT 协议开源。
推荐理由
原文给出 DeepSeek-V4.1-Flash 的参数结构、上下文长度和开源协议,读者可据此评估其部署与成本价值。
正文
@deepseek_ai just made Flash the new Pro. 🤯
DeepSeek-V4.1-Flash is live on SiliconFlow — Day 0.
552B MoE — ~8B active during prefill, ~16B during decode.
⚡ Faster inference, higher throughput
👁️ Native vision
🧠 1M context window
📉 ~4× smaller KV cache footprint vs. V4 Flash
📖 MIT licensed
Run it on SiliconFlow — production-ready, high-throughput inference that keeps Flash fast at scale.
Flash in name. Flagship in performance.
Build with it → http://siliconflow.com
来源:SiliconFlow · x.com