返回全部动态

Gemini 3.8 文本转语音模型发布

原标题:Gemini 3.8 text-to-speech says hello

Google DeepMind News一手来源模型发布质量 74

AI 摘要

Google DeepMind 发布 Gemini 3.8 Flash TTS 和 Gemini 3.8 Flash-Lite TTS 两款文本转语音模型,前者面向创意角色设计与逐行表演控制,后者面向高吞吐、低成本的配音与语音代理场景。新模型支持自然语言生成定制音色、30 秒样本声音复刻、2000+ 预置音色及 100 多种语言,并配备同意验证、SynthID 水印和 C2PA 凭证等安全措施。官方称其在 Hume AI 语音设计基准上排名第一,并在 Voice Arena 盲测中于日语、巴西葡萄牙语、越南语等语言上领先。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Gemini 3.8 text-to-speech says hello Today, we’re introducing two new text-to-speech models to the Gemini family, transforming voice generation from static presets into a dynamic creative studio. These models enable creators, developers, and enterprises to create richer, more expressive audio experiences, while enabling improved user experiences in products like Gemini Notebook and Google Vids. - Gemini 3.8 Flash TTS: Built for deep creative direction and character design. Create entirely new vo


发布时间:2026-09-23 23:25
抓取时间:2026-09-23 23:34
来源机构:Google DeepMind
阅读原文deepmind.google