PulseAugur
EN
LIVE 09:23:56
ENTITY LVBench

LVBench

PulseAugur coverage of LVBench — every cluster mentioning LVBench across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
4
7 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
4
7 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 7 TOTAL
  1. TOOL · CL_193450 ·

    MERIT framework simplifies ultra-long video understanding

    Researchers have developed MERIT, a novel framework for understanding ultra-long videos that exceed practical processing limits for current multi-modal large language models. MERIT employs a two-stage approach, first co…

  2. RESEARCH · CL_167440 ·

    New frameworks boost MLLM long-video understanding by adaptive frame processing · 3 sources tracked

    Three new research papers introduce novel frameworks for enhancing the long-video understanding capabilities of multimodal large language models (MLLMs). These approaches aim to overcome the limitations of fixed context…

  3. TOOL · CL_152094 ·

    New VideoTreeSearch framework enables self-correcting agents for long video QA

    Researchers have introduced VideoTreeSearch (VTS), a novel framework designed to improve long-video question answering by treating the task as a self-correcting search over an adaptive temporal tree. Unlike previous met…

  4. TOOL · CL_135427 ·

    Goal-Driven Data Optimization speeds up multimodal AI training

    Researchers have developed a framework called Goal-Driven Data Optimization (GDO) to improve the efficiency of multimodal instruction tuning. GDO computes sample descriptors to create optimized training subsets tailored…

  5. TOOL · CL_128730 ·

    New DELTAVID framework boosts video LLMs' fine-grained perception

    Researchers have introduced DELTAVID, a novel framework designed to improve the fine-grained spatiotemporal perception capabilities of video multimodal large language models (Video MLLMs). This approach transforms the t…

  6. RESEARCH · CL_97982 ·

    OmniAgent uses active perception for efficient video understanding · 2 sources tracked

    Researchers have introduced OmniAgent, a novel omni-modal agent designed for video understanding that utilizes an iterative Observation-Thought-Action cycle based on Partially Observable Markov Decision Processes (POMDP…

  7. RESEARCH · CL_95864 ·

    New research enhances vision-language models for medical, retrieval, and robotics tasks

    Researchers are developing new methods to improve vision-language models (VLMs) across various domains. One paper introduces CoT-Mediate, a framework to assess how generated reasoning influences VLM predictions in medic…