跳到正文
@testingcatalog· @testingcatalog · X·· 2026-09-02精选AI 评分71
AI 导读

Anthropic 宣布推出 Claude Fable 5.1 和 Claude Mythos 5.1,并称其面向编码与知识工作。据 Testing Catalog 引述,Fable 5.1 在 Terminal-Bench-Science 0.1 上得分 52.6%,是 Fable 5 的两倍多;在 Terminal-Bench 4.0 上得分 55.8%,Fable 5 为 42.0%。两款模型目前正在 Claude 上推送。

推荐理由

两款新模型在 Terminal-Bench 系列基准上的分数对比,可直观看出相对前代的提升幅度与推送状态。

正文

BREAKING 🔥: Anthropic has announced Claude Fable 5.1 and Claude Mythos 5.1!

> It scores 52.6% on Terminal-Bench-Science 0.1, more than double Fable 5.

> On Terminal-Bench 4.0, it scores 55.8% against 42.0% for Fable 5.

Rolling out on Claude now 👀 https://t.co/hxIt2yHzud https://t.co/SsGvoiWNvh

引用@claudeai@claudeai
We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work. https://t.co/8P9PSrWPi3
在 X 查看被引用的帖子

来源:@testingcatalog · x.com