跳到正文
原文
Google DeepMind·· 2026-04-16精选AI 评分60

Google DeepMind 发布 Gemini 3.1 Flash TTS 语音模型

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

AI 导读

Google DeepMind 发布 Gemini 3.1 Flash TTS,这是最新的文本转语音模型,支持 70 多种语言和原生多说话人对话。模型引入音频标签功能,可通过自然语言指令控制语速、语气和表达方式,在 Artificial Analysis TTS 排行榜上 Elo 得分 1,211。

推荐理由

官方介绍了新模型的音频标签控制能力和多语言覆盖,读者可据此评估它在语音生成应用中的可用性。

来源:Google DeepMind · deepmind.google