VGGSound
PulseAugur coverage of VGGSound — every cluster mentioning VGGSound across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
New AVENUE benchmark targets audio-video editing model evaluation
Researchers have introduced AVENUE, a new benchmark and evaluation framework designed to improve audio-video editing models. The AVENUE benchmark includes 1,291 source clips and 7,957 editing instructions, curated from …
-
AV-JEPA model advances audio-visual self-supervised learning
Researchers have introduced AV-JEPA, a new self-supervised learning model that extends LeJEPA to handle both audio and visual data. This model utilizes an early-fusion Vision Transformer and modality dropout for masking…
-
FoleyGenEx framework unifies video-to-audio generation with advanced controls
Researchers have introduced FoleyGenEx, a novel framework for unified video-to-audio generation that addresses limitations in existing methods. FoleyGenEx integrates multi-modal control, frame-level temporal alignment, …