Google DeepMind·· 2026-09-02精选AI 评分74
Google DeepMind 推出 Gemini 智能体视频理解功能
Introducing agentic video understanding with Gemini
AI 导读
Google DeepMind 在 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite 模型中推出智能体视频理解功能,该功能将模型核心推理与原生视频工具结合,动态搜索和检查目标视频片段,相比静态处理可降低最多 88% 的 token 消耗和 66% 的成本,同时提升最多 7% 的准确率。
推荐理由
原文给出了 token 消耗降低 88%、成本降低 66%、质量提升 7% 的具体数据,读者可以据此评估该功能对自身工作流的效率提升。
来源:Google DeepMind · deepmind.google