@Yuchenj_UW· @Yuchenj_UW · X·· 2026-08-26精选AI 评分77
AI 导读
GLM-5.3-Flash(Ox Alpha)发布,320B-A18B 规模不到 GLM-5.2 的一半,却在各项基准上全面超过 GLM-5.2。该模型原生多模态、支持 1M-token 上下文窗口,以 MIT 许可发布,此前以 Ox Alpha 名义预览并完全运行在中国 AI 芯片上,权重、API、Coding Plan、ZCode、Chat、AutoClaw 等官方入口已开放。Databricks 的 Yuchen Jin 表示将尽快把 GLM-5.3-Flash 提供给客户并让它跑得很快。
推荐理由
原文对比了 GLM-5.3-Flash 与 GLM-5.2 的参数量和基准成绩,可据此了解高效小模型的进展。
正文
GLM-5.3-Flash (Ox Alpha) is a big deal!
320B-A18B, less than half the size of GLM-5.2, yet it beats GLM-5.2 across every benchmark.
Amazing to see @Zai_org keep pushing frontier intelligence with increasingly efficient models. Appreciate the MIT license.
Databricks will bring GLM-5.3-Flash to our customers asap, and make it run super fast.
Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Previously previewed as Ox Alpha, running entirely on Chinese AI chips Blog: https://t.co/tzOmB7gdZP Available now across all official platforms: Weights: https://t.co/9LRMahY9Wa API: https://t.co/VcaQnzYmS9 Coding Plan: https://t.co/Nk8Y98HNhU ZCode: https://t.co/Peepqv4XSx Chat: https://t.co/WCqWT0qCQb AutoClaw: https://t.co/aGEG5HqTTb在 X 查看被引用的帖子
来源:@Yuchenj_UW · x.com