跳到正文
Hugging Face Daily Papers·· 2026-08-19AI 评分39

MoE-ViE:面向高效图像与视频理解的混合专家视觉编码器

MoE-ViE: Mixture of Experts Vision Encoder for Efficient Image and Video Understanding

阅读原文

本站未展示全文,请前往来源网站阅读。

AI 导读

研究团队提出 MoE-ViE 混合专家视觉编码器,系统探索 CLIP 风格视觉编码器的 MoE 设计,发现细粒度 MoE 拓扑显著优于稠密和标准 MoE 方案,并引入无辅助损失均衡变体与专用 MoE kernel 降低推理延迟。

来源:Hugging Face Daily Papers · arxiv.org