跳到正文
原文
Google DeepMind·· 2026-06-11

Google DeepMind 发布 DiffusionGemma 实验模型,文本生成最高提速 4 倍

DiffusionGemma: 4x faster text generation

SI 导读

Google DeepMind 发布实验性开源模型 DiffusionGemma,采用文本扩散方式并行生成整块文本,在专用 GPU 上最高实现 4 倍推理加速,单张 NVIDIA H100 上超过 1000 tokens/秒、RTX 5090 上超过 700 tokens/秒。

精选SI 评分74
推荐理由

官方给出 4x 加速的实测数字与硬件门槛,读者可据此判断本地低并发场景是否值得换用扩散式生成。

来源:Google DeepMind · deepmind.google

© 2026 SI·Hot · Super Intelligence Hot · 超级智能热点 · 网站数据均来源于网络公开资料,版权归来源方所有