@kimmonismus· @kimmonismus · X·· 2026-08-26精选AI 评分66
AI 导读
Zai 发布 GLM-5.3 Flash(代号 Ox Alpha),这是一款 320B MoE 模型,每 token 仅激活 18B 参数,开放权重、MIT 许可、原生多模态、1M 上下文。据 Zai 公布的数据,它在 Terminal-Bench 2.1 得 84.3,接近 Claude Opus 4.8 的 85.0;DeepSWE 得 63.4、AutomationBench 得 48.8,均为对比中的领先成绩,并在全部六项基准上超过更大的 GLM-5.2,而服务成本仅为后者的十分之一。原文同时提醒,18B 激活参数并不等于可本地运行的 18B 模型,全部 320B 权重仍需存储。
推荐理由
原文列出六项基准对比与 MIT 许可信息,读者可据此判断这一小激活参数模型的性价比。
正文
official: https://t.co/i7S70QQEsl
GLM-5.3 Flash ("Ox Alpha") official: Benchmarks attached. This looks exceptional for its size! GLM-5.3-Flash might be one of the most impressive efficiency releases yet. It is a 320B MoE with only 18B parameters active per token, yet Zai reports: - 84.3 on Terminal-Bench 2.1, nearly matching Claude Opus 4.8 at 85.0 - 63.4 on DeepSWE, ahead of Opus 4.8 and DeepSeek V4 Vision Exp - 48.8 on AutomationBench, ahead of Opus 4.8 and GPT-5.6 Terra - The highest GDPval-AA v2 score in its comparisonIt also beats the much larger GLM-5.2 across all six reported benchmarks while costing one-tenth as much to serve. Open weights, MIT licensed, natively multimodal, 1M context. Important caveat: 18B active parameters does not make it a normal local 18B model. All 320B weights still need to be stored. But in terms of intelligence per active parameter, this looks exceptional!在 X 查看被引用的帖子
来源:@kimmonismus · x.com