跳到正文
@dexhorthy· @dexhorthy · X·· 2026-09-06AI 评分23
AI 导读

SlopCodeBench 更新了多款模型在不同思考档位下的成绩:Fable 5.1 Medium 11%、GLM 5.3 Medium 7%、GPT-5.6 Sol medium 11%,GLM 5.3 High 19%、GPT-5.6 Sol XHigh 18%、GPT-6 Astra Xhigh 21% 仍在进行中。

正文

SlopCodeBench Update:

Fable 5.1 Medium - 11% - I think we need to re-run this at higher effort

GLM 5.3 Medium - 7%

GLM 5.3 High - 19%, still in progress

GPT-5.6 Sol medium - 11%

GPT-5.6 Sol XHigh - 18%, still in progress

GPT-6 Astra Xhigh - 21%, still in progress

This is the first time i've done a run with multiple models at different thinking settings, since I don't thinking effort translates cleanly across providers. My next fable reset is in 4 days so unless @trq212 can to find me some credits ❤️ I will have to hold off on the fable 5.1 xhigh until next week -

来源:@dexhorthy · x.com