跳到正文
@alibaba_cloud· @alibaba_cloud · X·· 2026-06-02AI 评分35
AI 导读

Performance2:多模态基准 Qwen3.7-Plus 的多模态提升并不局限于视觉理解上的孤立增益。相反,它们体现了多模态智能体所需核心能力的系统性增强——理解复杂视觉输入、对视觉信息进行推理、使用工具解决问题,并最终在代码或 GUI 环境中执行任务。

正文

Performance2:Multimodal Benchmarks

Qwen3.7-Plus’s multimodal improvements are not limited to isolated gains in visual understanding. Instead, they reflect a systematic enhancement of the core capabilities required by multimodal agents—understanding complex visual inputs, reasoning over visual information, using tools to solve problems, and ultimately executing tasks in code or GUI environments.

来源:@alibaba_cloud · x.com