返回全部动态
Gemini API 发布智能体视频理解,支持多模型并减少 token 消耗
原标题:September 1, 2026
AI 摘要
Gemini API 于 2026 年 9 月 1 日发布更新,为 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite 模型推出了智能体视频理解功能,该功能通过 Interactions 和 GenerateContent API 提供。模型能够动态导航视频时间线,按需请求转录、帧或音频轨道,与静态处理相比,长视频内容可减少高达 88% 的 token 使用量。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
Agentic video understanding: Released agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite across the Interactions and GenerateContent APIs. The model dynamically navigates video timelines, requesting transcripts, frames, or audio tracks on demand. This approach uses up to 88% fewer tokens for long-form content compared to static processing. To get started, see the Agentic video understanding guide.
发布时间:2026-09-01 08:00
抓取时间:2026-09-02 02:30
来源机构:Google