@testingcatalog· @testingcatalog · X·· 2026-09-02精选AI 评分71
AI 导读
Anthropic 宣布推出 Claude Fable 5.1 和 Claude Mythos 5.1,并称其面向编码与知识工作。据 Testing Catalog 引述,Fable 5.1 在 Terminal-Bench-Science 0.1 上得分 52.6%,是 Fable 5 的两倍多;在 Terminal-Bench 4.0 上得分 55.8%,Fable 5 为 42.0%。两款模型目前正在 Claude 上推送。
推荐理由
两款新模型在 Terminal-Bench 系列基准上的分数对比,可直观看出相对前代的提升幅度与推送状态。
正文
BREAKING 🔥: Anthropic has announced Claude Fable 5.1 and Claude Mythos 5.1!
> It scores 52.6% on Terminal-Bench-Science 0.1, more than double Fable 5.
> On Terminal-Bench 4.0, it scores 55.8% against 42.0% for Fable 5.
Rolling out on Claude now 👀 https://t.co/hxIt2yHzud https://t.co/SsGvoiWNvh
We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work. https://t.co/8P9PSrWPi3在 X 查看被引用的帖子
来源:@testingcatalog · x.com