AI 导读
OpenRouter 上线 GPT-6 Sol,定位为 Astra 之下、面向高要求编码与专业工作的档位。按 OpenAI 说法,它的事实性错误约为 GPT-5.6 Sol 的一半,DeepSWE v1.1 得分 68.8%,与 Claude Fable 5 的最佳成绩相差 1.1 分,单任务成本约低 80%。它还在 AutomationBench 上超过 Claude Opus 5,成本为后者的 9%。
推荐理由
文中给出 GPT-6 Sol 在编码基准上的得分与相对成本,可据此判断它对照 Claude 系列的性价比定位。
正文
GPT-6 Sol is the tier below Astra for demanding coding and professional work. Per OpenAI it makes about half as many factual mistakes as GPT-5.6 Sol, scores 68.8% on DeepSWE v1.1 (within 1.1 points of Claude Fable 5's best at ~80% lower cost per task), and beats Claude Opus 5 on AutomationBench at 9% of the cost per task.
Use it now: https://t.co/u4aFwZbOWn
来源:@OpenRouter · x.com