Gemini 3.8 文本转语音模型发布
原标题:Gemini 3.8 text-to-speech says hello
AI 摘要
Google DeepMind 发布 Gemini 3.8 Flash TTS 和 Gemini 3.8 Flash-Lite TTS 两款文本转语音模型,前者面向创意角色设计与逐行表演控制,后者面向高吞吐、低成本的配音与语音代理场景。新模型支持自然语言生成定制音色、30 秒样本声音复刻、2000+ 预置音色及 100 多种语言,并配备同意验证、SynthID 水印和 C2PA 凭证等安全措施。官方称其在 Hume AI 语音设计基准上排名第一,并在 Voice Arena 盲测中于日语、巴西葡萄牙语、越南语等语言上领先。
正文节选
Gemini 3.8 text-to-speech says hello Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio. These models enable creators, developers, and enterprises to create richer, more expressive audio experiences, while enabling improved user experiences in products like Gemini Notebook and Google Vids. - Gemini 3.8 Flash TTS: Built for deep creative direction and character design. Create entirely new vo