English(EN)KeyFrame-Compass: Towards Comprehensive Evaluation of Keyframe-Conditioned Video Generation
新的基准和框架应对视频生成一致性挑战 · 已追踪 4 个来源
作者PulseAugur 编辑部·[4 个来源]·
研究人员开发了新的基准和框架来评估和改进视频生成模型,特别关注多镜头和关键帧之间的视觉一致性。GroundShot 是一个新颖的框架,它使用实体级别的视觉记忆来在多镜头视频中保持一致性,而无需额外训练。KeyFrame-Compass 是一个全面的基准,旨在评估视频生成模型在保持整体视频质量的同时,对指定关键帧的遵循程度。此外,GeCo 是一个几何约束指标,用于检测并帮助减少生成视频中的几何变形和遮挡不一致问题。
AI
影响
这些评估指标和框架的进步对于推动 AI 驱动的视频生成能力的边界至关重要,能够实现更一致、视觉上更连贯的输出。
arXiv:2606.20799v2 Announce Type: replace-cross Abstract: Generating visually consistent multi-shot videos remains an open challenge. As videos span more shots, inconsistencies can accumulate across shots, causing entities that reappear across shots -- characters, objects, and lo…
Video generation increasingly relies on keyframe-based workflows, where creators specify a sequence of reference images to guide generation. Although recent models support multi-keyframe conditioning, it remains unclear whether they can faithfully reproduce the prescribed keyfram…
arXiv:2607.14202v1 Announce Type: new Abstract: Video generation increasingly relies on keyframe-based workflows, where creators specify a sequence of reference images to guide generation. Although recent models support multi-keyframe conditioning, it remains unclear whether they…
arXiv:2512.22274v3 Announce Type: replace Abstract: We introduce GeCo, a geometry-grounded metric for jointly detecting geometric deformation and occlusion-inconsistency artifacts in static scenes. By fusing residual motion and depth priors, GeCo produces interpretable, dense con…