跳到正文
原文
@MiniMax_AI· @minimax_ai · X·· 2026-06-04精选AI 评分69
AI 导读

MiniMax 官方宣布 M3 在 1M token 下解码速度提升 15.6 倍,并感谢 Fireworks AI 为其提供推理支持,用户可直接试用。其引用的 Fireworks AI 内容显示,M3 采用 MiniMax Sparse Attention(MSA),模型权重发布后也将在 Fireworks 社区提供。

推荐理由

官方给出 M3 在 1M token 下解码提速 15.6 倍,读者可据此判断其长上下文推理的工程取向。

正文

15.6× faster decoding at 1M tokens 🔥

Thanks @FireworksAI_HQ for powering the inference behind M3.

Try it now 👇

引用Fireworks AI (@FireworksAI_HQ)@FireworksAI_HQ
MiniMax M3 arrives with MiniMax Sparse Attention (MSA), 15.6x faster decoding at 1M tokens. We're partnering with @MiniMax_AI to power the inference behind this week's launch. Head to minimax.io to take it for a spin. Once the model weights are released, M3 will be available to the Fireworks community.
在 X 查看被引用的帖子

来源:@MiniMax_AI · x.com