@kimmonismus· @kimmonismus · X·· 2026-08-26精选AI 评分71
AI 导读
Qwen3.8-Flash-Next 正式发布,这是一个 125B 参数、另加 51B n-gram 嵌入、每 token 仅 6B 参数激活的稀疏 MoE 开源模型,在 9 项可比基准中有 8 项超过 Claude Opus 4.6 Max。
推荐理由
6B 激活参数的稀疏 MoE 在多项基准上对标 Claude Opus 4.6 Max,可据此比较开源小激活模型的能力位置。
正文
Qwen 3.8 Flash-Next official released: A 6B-active open model just beat Claude Opus 4.6 Max across 8 of 9 comparable benchmarks!
Qwen3.8-Flash-Next is a highly sparse MoE:
• 125B model parameters
• 51B additional n-gram embeddings
• Only 6B parameters active per token
It scores:
• 62.5 SWE-bench Pro
• 81.0 SWE-bench Multilingual
• 73.9 CoworkBench
• 55.7 JobBench
• 73.5 Toolathlon
• 81.3 IFBench
• 91.7 GPQA Diamond
• 91.9 LiveCodeBench
It also outperforms Qwen3.8-27B and DeepSeek-V4-Flash across most of the table.
Super cool release!!
来源:@kimmonismus · x.com