PulseAugur
实时 17:56:07
English(EN) We’re bringing agentic video understanding to our latest Gemini models.

Google DeepMind 为 Gemini 模型增强代理式视频理解能力

Google DeepMind 已为其 Gemini 模型引入了代理式视频理解能力,使其能够更有效地分析视频。Gemini 现在可以跨越文本记录、音频和帧进行推理,动态调整帧率以定位关键时刻,而不是处理整个文件。此增强功能对于长格式内容尤其有益,使用的 token 数量最多可减少 88%,在处理时长较长的录制内容时效率提升最为显著。该功能将通过 Google AI Studio 中的 API 逐步推广到 Gemini 3.7 Flash3.6 Flash3.5 Flash-Lite,并将很快在 Gemini 应用中提供。 AI

影响 提高了 AI 模型中视频分析的效率,有望降低长格式内容的成本并提高其性能。

排序理由 这是对现有模型的特性更新,并非新模型发布或重大的研究突破。

在 X — Google DeepMind 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Google DeepMind 为 Gemini 模型增强代理式视频理解能力

本文如何被排名

Signal score
22 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
这是对现有模型的特性更新,并非新模型发布或重大的研究突破。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
product, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [2]

  1. X — Google DeepMind TIER_1 English(EN) · GoogleDeepMind ·

    Gemini 不再扫描整个文件,而是跨越视频的文本、音频和帧进行推理,动态调整帧率以提取精确的时刻

    Instead of scanning an entire file, Gemini reasons across the video’s transcript, audio, and frames, dynamically adjusting the frame rate to pull the exact moments needed. The efficiency gains are most significant for long-form content, from 10-minute guides to multi-hour https:…

  2. X — Google DeepMind TIER_1 English(EN) · GoogleDeepMind ·

    我们将最新的Gemini模型引入了代理式视频理解功能。

    We’re bringing agentic video understanding to our latest Gemini models. They can now analyze videos with better accuracy while using up to 88% fewer tokens. 🧵 https://t.co/ZTjlaLk7tX