跳到正文
原文
Google DeepMind·· 2026-06-11精选AI 评分74

Google DeepMind 发布 DiffusionGemma 文本扩散模型,GPU 推理最高提速 4 倍

DiffusionGemma: 4x faster text generation

AI 导读

Google DeepMind 发布实验性开放模型 DiffusionGemma,采用文本扩散方式并行生成整块文本,在专用 GPU 上推理最高提速 4 倍。

推荐理由

原文给出 26B MoE 文本扩散模型的开放权重与硬件门槛,读者可据此判断本地低并发推理的取舍。

来源:Google DeepMind · deepmind.google