跳到正文
@kimmonismus· @kimmonismus · X·· 2026-08-21AI 评分28
AI 导读

社区猜测某新模型可能是 GLM-5.4/5.5,若属实将意味着模型能力在极短时间内再次大幅跃升。GLM-5.3 发布仅 7 天,相比 GLM-5.2 已有显著提升,且仅靠 RL 在相同基座模型上实现。作者认为这可能加速中美模型差距缩小,并迫使 OpenAI、Anthropic 加快发布节奏。

正文

If it's true that this is indeed GLM-5.4/5.5, then it would change everything, without exaggeration.

GLM-5.3 was released just 7 days ago and was an extremely significant leap compared to GLM-5.2, which was improved solely through real-time modeling (RL). Same base model. And all this in a very short time.

If it's true that GLM-5.4 has become so much better just a week later thanks to RL, it would demonstrate:

1) how much faster the models are now becoming. Not only is there no end in sight, but: now more than ever, exponential growth.

2) It would force OpenAI and Anthropic to release models. Anthropic, in particular, with its upcoming IPO, now has to prove itself. And it would put pressure on slowing down in favor of security.

3) And at least as importantly: the gap between China and the US is shrinking even further, even faster. It seems to be generally accepted that this is a Chinese model. No one suspects it's a Google model.

Given the same tokenizer, it's most likely GLM, and that would be the craziest thing we've seen in a long time for the reasons mentioned above.

But perhaps it's MiMo. Or, quite far-fetched, Ilya Sutskever's SSI model.

It remains exciting. I've rarely seen the community so impressed and confused at the same time. I'm equally confused and impressed.

@davis7 used Fable to determine which model it most closely resembles. A screenshot of the test is attached. h/t Ben Davis

来源:@kimmonismus · x.com