阿里云账号发布消息,Qwen-Image-2.1 现已原生支持透明图像生成,并且同样支持对透明图像进行编辑。推文给出了该能力的说明与示例图片,未披露更多技术细节或开放时间。



关注 AI 研究者、开发者与机构的动态
阿里云账号发布消息,Qwen-Image-2.1 现已原生支持透明图像生成,并且同样支持对透明图像进行编辑。推文给出了该能力的说明与示例图片,未披露更多技术细节或开放时间。



推荐理由:7B 轻量架构把生成与编辑统一到一个开放权重的模型中,可据此评估现有图像工作流的替换成本。
感谢 @sgl_project 的 day-0 支持!🙌 SGLang-Diffusion 现已支持 Qwen-Image-2.1:文生图、多图编辑,以及透明 RGBA 输出。快来试试!🎨
Day-0 support for @Alibaba_Qwen’s Qwen-Image 2.1 is here in SGLang-Diffusion! 🖥️ Native precision on a single RTX 4090 24GB with CPU offload - 1024×1024 generation in 18.7s and image editing in 21.7s with 22.7 GiB peak GPU memory during requests. - On an RTX PRO 6000 96GB: 8.0s generation and 9.6s editing. 🎨 Text-to-image, multi-image editing, and transparent RGBA output—all with one checkpoint. ⚡ Native inference with TP/SP, LoRA, and OpenAI-compatible APIs. 40 denoising steps, one image per request, warmed HTTP latency including PNG output. No quantization. Cookbook and GPU-specific commands below 👇
Qwen Image 2.1 is here! 🖼️ A 7B params native image generation and editing model, with up to 10 image references The model comes with it's own prompt enhancement LLMs, integrated with diffusers 🧨 and ComfyUI ▶️ on Spaces https://huggingface.co/spaces/hugging-apps/qwen-image-2-1
推荐理由:官方宣布生成与编辑共用单一 checkpoint,并提供浏览器免安装的 Spaces 演示,读者可直接上手体验。
阿易 AI Notes 把 Akshay 梳理的结构化判断模式提炼为 Choice 单选、Score 区间和布尔三种核心原语,用于智能体与后端系统中的分类与真假判断。
muse 做的一个有趣的资本配置游戏! 也可以叫另一个名字:“价格猜猜看:沃伦·巴菲特版” https://t.co/RBdjcCbBEI
Full video https://t.co/6E7I7dfiBI
下次你看到一条离谱到尴尬的假新闻标题时,记住这个 https://t.co/tvvI9EEj7M https://t.co/4t2OwXV5tF
啊啊啊 Tyler Cowen 喜欢 muse,这不是演习 @tylercowen https://t.co/ba91HdZAyh
阿易 AI Notes 拆解了 GitHub 开源项目 shhivv 旗下 third-hand 的三层结构,这是一套原生支持 macOS 的 0 截屏桌面自动化方案。
https://t.co/x8CzSSqhqk