返回全部动态

Gemini API 发布智能体视频理解,支持多模型并减少 token 消耗

原标题:September 1, 2026

Gemini API Release Notes一手来源API 更新质量 79

AI 摘要

Gemini API 于 2026 年 9 月 1 日发布更新,为 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite 模型推出了智能体视频理解功能,该功能通过 Interactions 和 GenerateContent API 提供。模型能够动态导航视频时间线,按需请求转录、帧或音频轨道,与静态处理相比,长视频内容可减少高达 88% 的 token 使用量。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Agentic video understanding: Released agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite across the Interactions and GenerateContent APIs. The model dynamically navigates video timelines, requesting transcripts, frames, or audio tracks on demand. This approach uses up to 88% fewer tokens for long-form content compared to static processing. To get started, see the Agentic video understanding guide.


发布时间:2026-09-01 08:00
抓取时间:2026-09-02 02:30
来源机构:Google
阅读原文ai.google.dev