Artificial Analysis 统计显示,随着 MiMo-V2.6-Pro、Claude Opus 5.5、GPT-6 Luna 和 GPT-6 Sol 本周发布,智能指数对单任务成本的帕累托前沿新增 11 个点,其中 GPT-6 Luna 贡献 5 个、Claude Opus 5.5 贡献 4 个,MiMo-V2.6-Pro 和 GPT-6 Sol 各贡献 1 个。
文中列出四款新模型在智能指数与单任务成本上的前沿点,可用来对照不同价格区间的模型表现。
The Intelligence Index vs Cost per Task Pareto frontier shifted this week with the releases of MiMo-V2.6-Pro, Claude Opus 5.5, GPT-6 Luna, and GPT-6 Sol
Together they have established eleven new points on the Pareto frontier (driven by different reasoning efforts): five from GPT-6 Luna, one each from MiMo-V2.6-Pro and GPT-6 Sol, and four from Claude Opus 5.5.
GPT-6 Luna (max) scores 37 at $0.068 per task, MiMo-V2.6-Pro scores 46 at $0.13, GPT-6 Sol (max) scores 48 at $1.06, and Claude Opus 5.5 (max with fallback) is the new highest-scoring model at 58 at $5.98.
来源:@ArtificialAnlys · x.com