跳到正文
@kimmonismus· @kimmonismus · X·· 2026-05-27AI 评分31
AI 导读

看到整体风向转向支持 Codex,真是令人惊叹。 我读到很多帖子说,有了 GPT-5.5,Codex 现在真的很强,而 Claude Code 也经常更受青睐。 (我自己已经成了 Codex 的超级粉丝)。 与此同时,新的 DeepSWE 基准显示,GPT-5.5 在这项评测中也排名第一。

正文

It's truly amazing to see how the general sentiment has shifted in favor of Codex.

I'm reading so many posts saying that Codex is really good now with GPT-5.5, and that Claude Code is regularly preferred.

(I've become a huge Codex fan myself).

At the same time, the new DeepSWE benchmark shows that GPT-5.5 is now ranked number one in this measurement as well.

引用Serena Ge (Datacurve) (@serenaa_ge)@serenaa_ge
Today we’re releasing DeepSWE, a new standard for agentic coding benchmarks. On public leaderboards, top models often look relatively close in capability. DeepSWE shows where they actually diverge, reflecting the realistic experience of developers in their day-to-day work.
在 X 查看被引用的帖子

来源:@kimmonismus · x.com