PulseAugur
实时 09:27:12
English(EN) Long Story Short: Story-level Video Understanding from 20K Short Films

新的 SF20K 数据集通过 20,000 部业余电影推进视频理解

研究人员推出了 Short-Films 20K (SF20K),这是一个包含 20,000 多部业余电影的新数据集,总时长达 3,582 小时,旨在推动视频理解超越简短、范围有限的片段。该数据集旨在解决现有视频数据集的局限性,例如叙事狭窄和潜在的数据泄露。SF20K 配备了 SF20K-Test,这是一个包含 95 部电影和近 1,000 个问答对的问答基准测试,证明了视频理解模型进行长期推理的必要性。 AI

影响 该数据集可以使视频理解模型实现更复杂的长期推理,从而可能改进内容分析和摘要等应用。

排序理由 该条目描述了一个新的学术视频理解数据集和基准测试。[lever_c_demoted from research: ic=1 ai=1.0]

在 arXiv cs.CL 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的 SF20K 数据集通过 20,000 部业余电影推进视频理解

本文如何被排名

Signal score
13 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该条目描述了一个新的学术视频理解数据集和基准测试。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
Breaking (< 6h)
Fresh story with cross-source coverage still developing. Ranking may shift as more sources report.

完整方法见我们的编辑标准

报道来源 [1]

  1. arXiv cs.CL TIER_1 English(EN) · Ridouane Ghermi, Xi Wang, Vicky Kalogeiton, Ivan Laptev ·

    长话短说:从 20K 短片进行故事级视频理解

    arXiv:2406.10221v3 Announce Type: replace-cross Abstract: Recent developments in vision-language models have significantly advanced video understanding. Existing datasets and tasks, however, have notable limitations. Most datasets are confined to short videos with limited events …