inclusionAI 开源 ZwZ 系列模型:用 Region-to-Image 蒸馏实现「不缩放也能缩放」的细粒度多模态感知
inclusionAI/Zooming-without-Zooming
阅读原文
本站未展示全文,请前往来源网站阅读。
AI 导读
inclusionAI 开源 ZwZ 系列多模态模型(2B/4B/7B/8B),基于 Qwen3-VL 与 Qwen2.5-VL,通过 Region-to-Image Distillation 把推理时的区域缩放转为训练时蒸馏,单次前向即可完成细粒度感知,在开源模型中达到 SOTA。
来源:inclusionAI GitHub 新仓库 · github.com