PulseAugur
实时 07:08:53
English(EN) ScanFocus: A Coarse-to-Fine Framework for Spatio-Temporal Video Grounding

ScanFocus框架通过粗粒度到细粒度方法增强视频定位

研究人员推出ScanFocus,一个旨在改进时空视频定位(STVG)的新框架。该方法解决了在视频中平衡全局上下文与精确对象定位的挑战,而当前方法由于计算成本和时间下采样常常难以应对。ScanFocus采用粗粒度到细粒度策略,首先进行全局扫描生成初始提案,然后通过语义引导时间聚合器(SGTA)进行细化,以捕捉细粒度细节和快速运动变化,从而实现精确的时间戳回归。在多个基准测试上的实验表明,ScanFocus的性能优于现有方法。 AI

影响 这个新框架可以提高依赖于理解视频内容中对象轨迹的视频分析系统的准确性和效率。

排序理由 该集群包含一篇详细介绍特定AI任务新框架的研究论文。

在 arXiv cs.AI 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

ScanFocus框架通过粗粒度到细粒度方法增强视频定位

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群包含一篇详细介绍特定AI任务新框架的研究论文。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
42 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准

报道来源 [2]

  1. arXiv cs.AI TIER_1 English(EN) · Kai Chen, Ming Dai, Wenxuan Cheng, Wankou Yang ·

    ScanFocus:用于时空视频基础的粗到细框架

    arXiv:2607.13421v1 Announce Type: cross Abstract: Spatio-Temporal Video Grounding (STVG) aims to retrieve the visual trajectory of a specific object from a video stream as described by a natural language expression. However, most advanced methods struggle to balance global contex…

  2. arXiv cs.AI TIER_1 English(EN) · Wankou Yang ·

    ScanFocus:一种用于时空视频基础的粗粒度到细粒度框架

    Spatio-Temporal Video Grounding (STVG) aims to retrieve the visual trajectory of a specific object from a video stream as described by a natural language expression. However, most advanced methods struggle to balance global context modeling with precise boundary localization. Due…