Google DeepMind·· 2026-04-16精选AI 评分60
Google DeepMind 发布 Gemini 3.1 Flash TTS 语音模型
Gemini 3.1 Flash TTS: the next generation of expressive AI speech
AI 导读
Google DeepMind 发布 Gemini 3.1 Flash TTS,这是最新的文本转语音模型,支持 70 多种语言和原生多说话人对话。模型引入音频标签功能,可通过自然语言指令控制语速、语气和表达方式,在 Artificial Analysis TTS 排行榜上 Elo 得分 1,211。
推荐理由
官方介绍了新模型的音频标签控制能力和多语言覆盖,读者可据此评估它在语音生成应用中的可用性。
来源:Google DeepMind · deepmind.google