AI 导读
ComfyUI 在5月集成了11个覆盖图像、3D、音频、视频和多模态的新模型,全部作为可直接拖拽的节点上线。
正文
记得4月我们内部Apple 给大家介绍ComfyUI工作流时,很多人一脸懵逼!
因为他们平时都是玩豆包、DeepSeek 的!
压根没有接触过ComfyUI 这玩意,但是也和客户,以及周围干业务的人接触知道。
其实这玩意在工作流生产环节中使用的频率非常高!
我也发现一个趋势越来越明…
5月他们悄无声息地集成了11个跨图像、3D、音频、视频和多模态的新模型。
最亮眼的几个直接可以把项目效率拉高了一个量级。
Krea 2 把风格优先的图像生成直接拉进来,第一天就以Partner Node形式上线。
它不再只拼画面里有什么,是把整个画面的感觉做到极致。
VOID来自Netflix,能把对象连同它带来的阴影、反射、物理交互全部干净移除,Apache 2.0开源,原生支持。
Tripo 3.1加TripoSplat,则实现了一张图直接出完整3D Gaussian资产,全流程端到端。
此外Gemma 4、Stable Audio 3、BiRefNet、MoGe、Claude、OpenRouter、Luma UNI-1也同步上线。
这些模型以前可能还得单独开云端账号、调API、处理格式兼容。
现在全变成ComfyUI里的节点,随手拖拽就能串成复杂工作流。
这其实戳破了一个共识:AI进步不是靠单一模型越来越大,而是靠本地工具把最新能力快速变成可组合、可重复的生产力。
ComfyUI把前沿研究直接转化成每个人都能本地跑的节点,真正让创作者把控制权握在自己手里。
Video
In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generation, live as a Partner Node on day one. Competes on how the frame looks, not just what's in it → VOID (Netflix) — removes objects and everything they caused: shadows, reflections, physical interactions. Apache 2.0, natively supported → Tripo 3.1 + TripoSplat — one image to a full 3D Gaussian asset, end to end Also live in May: Gemma 4, HidDream-O1-Image, Stable Audio 3, BiRefNet, MoGe, Claude, OpenRouter, Luma UNI-1 Learn more with our blog 👇 Video在 X 查看被引用的帖子
来源:@berryxia · x.com