AI 导读
有人在消费级 GPU 上跑出了 Opus 4.6 级别的智能。 Qwen3.8-27B-GGUF 在 2×4090 上达到 80 tok/s,完整 262,144 上下文,开启 MTP,仅占用 48GB 显存中的 34GB。 注意,Qwen3.8-27B:SWE-Pro 61.7 对 Opus 53.4;GPQA 89.2 对 91.3。 https://t.co/2YNCUBPX0I
正文
Somebody running an Opus 4.6-level intelligence on consumer GPU.
Qwen3.8-27B-GGUF doing 80 tok/s on 2×4090s, full 262,144 context, MTP on, and only 34GB of 48GB VRAM occupied.
Note, Qwen3.8-27B: SWE-Pro 61.7 vs Opus 53.4; GPQA 89.2 vs 91.3.
https://t.co/2YNCUBPX0I
来源:@rohanpaul_ai · x.com