SsV2
PulseAugur coverage of SsV2 — every cluster mentioning SsV2 across labs, papers, and developer communities, ranked by signal.
-
COMET framework boosts video LLMs with enhanced motion and temporal reasoning
Researchers have developed COMET, a new framework designed to enhance video multimodal large language models by improving their understanding of fine-grained motion and temporal reasoning. The framework introduces a ded…
-
New Mamba-like attention model VideoSEMA boosts video understanding performance
Researchers have introduced VideoSEMA, a novel Mamba-like attention model designed for efficient and scalable video understanding. This model utilizes a split space-time attention mechanism, combining local window atten…
-
HyperGS advances video representation with fast, generalizable Gaussian predictions
Researchers have developed HyperGS, a novel feedforward approach that directly predicts Gaussian representations for videos in a single pass, eliminating the need for per-video optimization. This method significantly sp…
-
HyperGS advances Gaussian video representation with faster, generalizable predictions
Researchers have developed HyperGS, a novel feedforward approach for video representation using Gaussian Splatting. Unlike previous methods requiring per-video optimization, HyperGS directly predicts Gaussian representa…
-
Grounding Video Reasoning in Physical Signals
Researchers have developed a new benchmark for evaluating physical video understanding, moving beyond simple event recognition to assess a model's ability to pinpoint events in time and space. This benchmark, which incl…