跳到正文
原文
Google DeepMind·· 2026-04-16精选AI 评分60

Google DeepMind 发布 Gemini 3.1 Flash TTS 语音模型

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

AI 导读

Google DeepMind 发布 Gemini 3.1 Flash TTS 文本转语音模型,通过在文本中嵌入自然语言 audio tags 实现对语音风格、语速和表达的细粒度控制。

推荐理由

原文给出了基准分数和开放入口,读者可据此判断其在语音生成上的定位与可用性。

来源:Google DeepMind · deepmind.google