AI 导读
SlopCodeBench 更新了多款模型在不同思考档位下的成绩:Fable 5.1 Medium 11%、GLM 5.3 Medium 7%、GPT-5.6 Sol medium 11%,GLM 5.3 High 19%、GPT-5.6 Sol XHigh 18%、GPT-6 Astra Xhigh 21% 仍在进行中。
正文
SlopCodeBench Update:
Fable 5.1 Medium - 11% - I think we need to re-run this at higher effort
GLM 5.3 Medium - 7%
GLM 5.3 High - 19%, still in progress
GPT-5.6 Sol medium - 11%
GPT-5.6 Sol XHigh - 18%, still in progress
GPT-6 Astra Xhigh - 21%, still in progress
This is the first time i've done a run with multiple models at different thinking settings, since I don't thinking effort translates cleanly across providers. My next fable reset is in 4 days so unless @trq212 can to find me some credits ❤️ I will have to hold off on the fable 5.1 xhigh until next week -
来源:@dexhorthy · x.com