跳到正文
@googleaidevs· @googleaidevs · X·· 2026-06-11AI 评分64
AI 导读

Google 发布实验性开源模型 DiffusionGemma,采用 Apache 2.0 许可,探索文本扩散生成,官方称在专用 GPU 上 token 输出最高快 4 倍。

正文

DiffusionGemma, our experimental open model released under an Apache 2.0 license, explores text diffusion, an exceptionally fast approach to text generation.

Here’s how DiffusionGemma accelerates development:

+ Faster token output: By shifting the bottleneck from memory bandwidth to raw compute, the model generates up to 4x faster token output on dedicated GPUs

+ Accessible hardware footprint: Activates just 3.8B parameters during inference, fitting comfortably within 24GB-VRAM high-end consumer GPUs when quantized

+ Novel workflows: Parallel token generation enables self-correction, making it ideal for code infilling, in-line editing, and non-linear structures

DiffusionGemma prioritizes speed over raw quality and accelerates best on compute-bound hardware (like @NVIDIAAI GPUs). Standard @GoogleGemma 4 remains recommended for production quality and memory-bound devices.

来源:@googleaidevs · x.com