跳到正文
@emollick· @emollick · X·· 2026-08-19AI 评分32
AI 导读

Qwen 27B 确实是个很不错的本地模型,但用起来就会发现,它在智能体任务上明显远不如这里列出的其他模型,尤其是在 GDPval-AA 声称要衡量的那类复杂任务上。 自己动手做基准测试吧!https://t.co/r2p88zJTfr

正文

Qwen 27B is really good local model but, when you use it, it is immediately absolutely and obviously nowhere near as good as the other models listed here for agentic tasks, and especially for the kinds of complex tasks that GDPval-AA proports to measure

Do your own benchmarking! https://t.co/r2p88zJTfr

引用@kimmonismus@kimmonismus
Qwen 27B is the "DeepSeek moment" for open source. It matches the closed-source state-of-the-art from just a few months ago and runs on an RTX 5090. Without exaggeration, it’s a game changer. https://t.co/1MYBuc70KQ https://t.co/eS6bxDxDu2
在 X 查看被引用的帖子

来源:@emollick · x.com