返回全部动态

Google 发布 Flash TTS 系列模型,支持文本描述设计音色

原标题:Google's new Flash TTS models let you design AI voices from scratch using text descriptions

THE DECODER模型发布质量 71

AI 摘要

Google 发布两款文本转语音模型 Gemini 3.8 Flash TTS 和 Flash-Lite TTS,支持 100 多种语言。Flash TTS 可通过文本描述从零设计音色,并支持用 30 秒音频样本进行声音克隆,Flash-Lite TTS 面向低成本大规模语音生成。两款模型均支持逐行舞台指示、双人对话和非语言音效,通过 Gemini API 和 Google AI Studio 推出,企业版 API 随后跟进。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Google's new Flash TTS models let you design AI voices from scratch using text descriptions Key Points - Google has released two new text-to-speech models, Gemini 3.8 Flash TTS and Flash-Lite TTS, which support more than 100 languages. Flash TTS can create new voices from text descriptions, and a voice cloning feature builds voice profiles from 30-second audio samples. - Flash TTS is aimed at creative uses such as podcasts, audiobooks, and game characters, while Flash-Lite TTS is designed for lo


发布时间:2026-09-24 01:39
抓取时间:2026-09-24 02:16
来源机构:THE DECODER
阅读原文the-decoder.com