跳到正文
@testingcatalog· @testingcatalog · X·· 20 天前AI 评分61
AI 导读

阿里发布全模态模型 Qwen3.8-Omni-Flash,支持音频与视频理解,具备 1M token 上下文窗口,在 WildClawBench-MM 和 UniClawBench 上的表现接近 Gemini 3.8 Flash。随附图片将其定位为 Native Agent,称其先理解查询,再决定如何听或看,无需看完整段视频。

正文

Alibaba released Qwen3.8-Omni-Flash, a new omni-modal model with audio and video understanding.

Qwen3.8-Omni-Flash comes with 1M-token context window and performance close to Gemini 3.8 Flash on WildClawBench-MM & UniClawBench. https://t.co/jata5ZHXcE

来源:@testingcatalog · x.com