ActivityNet: A large-scale video benchmark for human activity understanding
PulseAugur coverage of ActivityNet: A large-scale video benchmark for human activity understanding — every cluster mentioning ActivityNet: A large-scale video benchmark for human activity understanding across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New Event ActivityNet benchmark advances untrimmed action understanding
Researchers have introduced Event ActivityNet, a new benchmark designed to advance the field of untrimmed action understanding in videos. This dataset, derived from the existing ActivityNet videos, features over 3,200 v…
-
New PHA-Net improves text-video retrieval with prototype alignment
Researchers have developed PHA-Net, a novel network for text-video retrieval that utilizes shared prototypes to align cross-modal representations efficiently. This approach addresses the semantic mismatch between text a…
-
MLLMs show positional bias in multi-video summarization
Researchers have identified a positional bias in multimodal large language models (MLLMs) when summarizing multiple videos. This bias means the quality of a summary can depend on the order in which videos are presented …