跳到正文
原文
Google DeepMind·· 2026-09-02精选AI 评分78

Google DeepMind 在 Gemini 中推出智能体视频理解功能

Introducing agentic video understanding with Gemini

AI 导读

Google DeepMind 宣布在 Gemini 3.7 Flash、3.6 Flash 和 3.5 Flash-Lite 模型中推出智能体视频理解功能。该功能通过让模型主动决定观看内容、速度和模态来替代固定帧率的静态处理,使 token 消耗最多降低 88%,成本最多减少 66%,准确率提升最高达 7%。

推荐理由

官方给出了视频理解从静态处理转向智能体动态检索的技术路径,并提供了具体的成本与精度优化数据。

来源:Google DeepMind · deepmind.google