跳到正文
Google DeepMind·· 2026-06-11精选AI 评分79

Google DeepMind 发布文本扩散模型 DiffusionGemma,GPU 推理速度最高提升 4 倍

DiffusionGemma: 4x faster text generation

AI 导读

Google DeepMind 推出基于文本扩散机制的开源实验模型 DiffusionGemma,采用 26B MoE 架构且推理仅激活 3.8B 参数,在专用 GPU 上实现最高 4 倍的文本生成速度。

推荐理由

原文给出了文本扩散架构在单卡低并发推理下的速度收益与质量妥协,读者可以据此评估它在实时代码补全或本地交互场景的应用可行性。

来源:Google DeepMind · deepmind.google