跳到正文
原文
Google DeepMind·· 2026-04-16精选AI 评分72

Google DeepMind 推出 Gemini 3.1 Flash TTS 语音模型

Gemini 3.1 Flash TTS: the next generation of expressive AI speech

AI 导读

Google DeepMind 正式推出文本转语音模型 Gemini 3.1 Flash TTS,支持通过自然语言音频标签精确控制语音风格、节奏和表达,并原生支持多说话人对话与 70 多种语言。

推荐理由

该模型通过自然语言行内标签实现对语调和语速的精细调度,为构建多角色对话系统提供了更具成本效益的语音方案。

来源:Google DeepMind · deepmind.google