跳到正文
原文
@kilocode· @kilocode · X·· 2026-06-02精选AI 评分67
AI 导读

MiniMax 发布 M3 模型,官方称其为首个同时整合编程、Agent 与多模态三项前沿能力的开放权重模型。官方报告其在 SWE-Bench Pro 取得 59.0%、Terminal Bench 2.1 取得 66.0%、SWE-fficiency 取得 34.8%、KernelBench Hard 取得 28.8%、MCP Atlas 取得 74.2%,并通过 MiniMax Sparse Attention 将上下文扩展到 1M。模型从第一步起原生多模态,权重与技术报告预计约 10 天后发布,API 与 MiniMax Code 入口已开放。

推荐理由

官方给出 M3 的编程与 Agent 基准数据及 1M 上下文能力,可据此观察开源权重模型的前沿能力边界。

正文

MiniMax M3 is live, and the benchmarks are turning heads.

MiniMax reports 59.0% on SWE-Bench Pro, 66.0% on Terminal Bench 2.1, and 74.2% on MCP Atlas, with sparse attention scaling context to 1M tokens.

Congrats to the MiniMax team on the Day 0 launch! 🚀

引用MiniMax (official) (@MiniMax_AI)@MiniMax_AI
Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency, 28.8% KernelBench Hard, 74.2% MCP Atlas - MiniMax Sparse Attention scales context to 1M - Natively Multimodal from Step Zero API: platform.minimax.io Token Plan: platform.minimax.io/subscrib… 🚀New! MiniMax Code: code.minimax.io Weights & Tech Report in ~10 Days
在 X 查看被引用的帖子

来源:@kilocode · x.com