跳到正文
@alexandr_wang· @alexandr_wang · X·· 2026-09-03AI 评分54
AI 导读

Meta 的 Muse Spark 1.3 (max) 在 Artificial Analysis 编码智能体指数中于 Muse Code 框架下得到 68 分,仅次于 Claude Opus 5 (xhigh)。已开放的 Muse Spark 1.3 (xhigh) 得 64 分,较 8 月的 1.2 版提升 2 分,单任务成本 1.72 美元,为 60 分以上智能体中最低。max 版处于有限预览阶段,Meta 尚未公布定价。

正文

muse spark 1.3 + muse code evals competitively with Claude code + opus 5 and Claude code + fable 5 https://t.co/rfYfjfQLJO

引用@ArtificialAnlys@ArtificialAnlys
Meta's Muse Spark 1.3 (max), which is in limited preview for Meta's partners, scores 68 on the Artificial Analysis Coding Agent Index in the Muse Code harness, #2 behind only Claude Opus 5 (xhigh) in Claude Code. The variant available now, Muse Spark 1.3 (xhigh), scores 64 and costs the least per task of any agent above a 60 index score Muse Spark 1.3 (xhigh) enters the Artificial Analysis Coding Agent Index at 64 in Muse Code, up 2 points from Muse Spark 1.2 (62, August). It enters level with Grok 4.5 (high) in Grok Build (64) and behind GPT-5.6 Sol (max) in Codex (65). At $1.72 per task, it costs the least of any agent above a 60 index score, around a fifth of the cost of Claude Opus 5 (xhigh) in Claude Code ($8.17) Muse Spark 1.3 (max), which is in a limited preview stage, lands at 68 in Muse Code. It enters behind only Claude Opus 5 (xhigh) in Claude Code (68), and ahead of Claude Fable 5 (max) in Claude Code (67) and GPT-5.6 Sol (max) in Codex (65). Muse Spark 1.3 (max) is excluded from cost comparisons as Meta has not announced pricing for the limited release Claude Fable 5.1 results are in progress and will be added when complete. Congratulations @AIatMeta, @finkd, and @alexandr_wang on this result!
在 X 查看被引用的帖子

来源:@alexandr_wang · x.com