AI 导读
用户实测 MiniMax M3 前端能力后给出正面评价,称其性价比高、不偷懒且会充分推理,能在多个设计方案间更好地权衡。该模型由 MiniMax 发布,被称为首个融合三项前沿能力的开放权重模型,在 SWE-Bench Pro 取得 59.0%、Terminal Bench 2.1 取得 66.0%,并通过 MiniMax Sparse Attention 将上下文扩展至 1M。实测者表示下一步将测试后端任务。
正文
MiniMax M3’s frontend capabilities are pretty nice
very strong model for the price.
not lazy, thinks through the task (thinks a lot), and doesn’t just take the shortest path
M3 can reason between multiple design choices better than i expected
with the right skills around it, m3 seems strong model
will test backend tasks next.
Video
Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Bench Pro, 66.0% Terminal Bench 2.1, 34.8% SWE-fficiency, 28.8% KernelBench Hard, 74.2% MCP Atlas - MiniMax Sparse Attention scales context to 1M - Natively Multimodal from Step Zero API: platform.minimax.io Token Plan: platform.minimax.io/subscrib… 🚀New! MiniMax Code: code.minimax.io Weights & Tech Report in ~10 Days在 X 查看被引用的帖子
来源:@lostinlatencyX · x.com