Google DeepMind·· 2026-06-11精选AI 评分74
Google DeepMind 开源实验性扩散文本模型 DiffusionGemma,推理提速最高 4 倍
DiffusionGemma: 4x faster text generation
AI 导读
Google DeepMind 发布实验性开源模型 DiffusionGemma,采用文本扩散方式一次并行生成整段文本,在专用 GPU 上文本生成速度最高提升 4 倍,单张 NVIDIA H100 可达 1000+ tokens per second,NVIDIA GeForce RTX 5090 可达 700+ tokens per second。
推荐理由
原文给出了速度、显存与质量取舍的具体数据,读者可据此判断它是否适合本地交互场景。
来源:Google DeepMind · deepmind.google