跳到正文
@ArtificialAnlys· @ArtificialAnlys · X·· 13 天前精选AI 评分66
AI 导读

Artificial Analysis 统计显示,随着 MiMo-V2.6-Pro、Claude Opus 5.5、GPT-6 Luna 和 GPT-6 Sol 本周发布,智能指数对单任务成本的帕累托前沿新增 11 个点,其中 GPT-6 Luna 贡献 5 个、Claude Opus 5.5 贡献 4 个,MiMo-V2.6-Pro 和 GPT-6 Sol 各贡献 1 个。

推荐理由

文中列出四款新模型在智能指数与单任务成本上的前沿点,可用来对照不同价格区间的模型表现。

正文

The Intelligence Index vs Cost per Task Pareto frontier shifted this week with the releases of MiMo-V2.6-Pro, Claude Opus 5.5, GPT-6 Luna, and GPT-6 Sol

Together they have established eleven new points on the Pareto frontier (driven by different reasoning efforts): five from GPT-6 Luna, one each from MiMo-V2.6-Pro and GPT-6 Sol, and four from Claude Opus 5.5.

GPT-6 Luna (max) scores 37 at $0.068 per task, MiMo-V2.6-Pro scores 46 at $0.13, GPT-6 Sol (max) scores 48 at $1.06, and Claude Opus 5.5 (max with fallback) is the new highest-scoring model at 58 at $5.98.

来源:@ArtificialAnlys · x.com