V Star
PulseAugur coverage of V Star — every cluster mentioning V Star across labs, papers, and developer communities, ranked by signal.
-
New frameworks decouple AI perception and reasoning for enhanced visual understanding · 6 sources tracked
Researchers have introduced novel frameworks to enhance the fine-grained visual reasoning capabilities of vision-language models. Rule-VLN addresses the challenge of embodied AI agents prioritizing physical navigation o…
-
New SER method enhances Video MLLM reasoning with semantic evidence rewards · 4 sources tracked
Researchers have developed a new method called Semantic Evidence Reward (SER) to improve the spatio-temporal reasoning capabilities of Video Multimodal Large Language Models (Video MLLMs). Existing models often struggle…
-
Grounding Video Reasoning in Physical Signals
Researchers have developed a new benchmark for evaluating physical video understanding, moving beyond simple event recognition to assess a model's ability to pinpoint events in time and space. This benchmark, which incl…