Google DeepMind 发布 Gemini 3.8 Live 系列语音模型
原标题:Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
AI 摘要
Google DeepMind 发布 Gemini 3.8 Live 与 Gemini 3.8 Live Extended Thinking 两款面向近实时语音交互的新模型。前者主打规模化与成本效率,支持近实时视觉输入、97 种语言自动切换及后台工具调用;后者面向高复杂度任务,可边推理边说话,在 Artificial Analysis 语音到语音质量指数上以 82.6 分位列第一,并在 τ-Voice 等智能体任务基准上领先。两款模型即日起通过 Gemini API、Google AI Studio 等渠道向开发者、企业和普通用户逐步开放,所有生成音频均嵌入 SynthID 水印。
正文节选
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. - Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. - Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-