跳到正文
Gemini API:更新日志·· 2026-09-01精选AI 评分61

Gemini API 为 Gemini 3.7 Flash 等模型推出 Agentic 视频理解

September 1, 2026

AI 导读

Gemini API 为 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite 推出 Agentic 视频理解,覆盖 Interactions 和 GenerateContent 两个 API。模型可动态导航视频时间线,按需请求转写、帧或音频轨,长视频内容相比静态处理最多节省 88% 的 token。

推荐理由

官方说明动态按需取用转写、帧和音频的机制,读者可据此评估长视频处理的 token 成本变化。

正文 · 原文
  • Agentic video understanding: Released agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite across the Interactions and GenerateContent APIs. The model dynamically navigates video timelines, requesting transcripts, frames, or audio tracks on demand. This approach uses up to 88% fewer tokens for long-form content compared to static processing.

    To get started, see the Agentic video understanding guide.

来源:Gemini API:更新日志 · ai.google.dev