VBench-Long
PulseAugur coverage of VBench-Long — every cluster mentioning VBench-Long across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
MLLMs Enhance Text-to-Video Generation with Semantic Correction and Visual Planning
Two new research papers explore enhancing text-to-video generation by integrating multimodal large language models (MLLMs) with diffusion models. The first paper introduces a framework that injects MLLM feedback directl…
-
Alaya-EVOKE introduces interactive world model for open-ended video generation
Researchers have introduced Alaya-EVOKE, an interactive world model designed for open-ended video generation. The model utilizes external persistent memory and a novel long-horizon teacher to achieve responsive generati…
-
New framework enhances long video generation with adaptive resource allocation
Researchers have developed a new framework called Surprise Forcing to improve the generation of long videos by diffusion models. This method addresses limitations in current streaming autoregressive diffusion models, wh…
-
New methods enhance autoregressive video generation quality and efficiency
Researchers are developing new methods to improve autoregressive video generation, focusing on efficiency and quality. One approach, One-Forcing, combines a DMD objective with a GAN loss to achieve stable, high-quality …
-
Pyramid Forcing improves long video generation with head-aware cache policy
Researchers have introduced Pyramid Forcing, a novel KV cache policy designed to enhance the quality of long video generation. This method addresses the issue of accumulated errors in autoregressive video synthesis by r…