各家基准对 GPT-6 Astra 评价不一,ARC-AGI-3 效率超人类均值促使 Chollet 上调 AGI 预测
Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward
阅读原文
本站未展示全文,请前往来源网站阅读。
AI 导读
针对 OpenAI 的 GPT-6 Astra,各家基准给出矛盾结论,Epoch AI 给出 169 分使其领先,Artificial Analysis 则认为它不优于前代、落后于 Claude Fable 5.1。
来源:The Decoder · the-decoder.com