PulseAugur
EN
LIVE 11:38:29
ENTITY Large Vision-Language Model

Large Vision-Language Model

PulseAugur coverage of Large Vision-Language Model — every cluster mentioning Large Vision-Language Model across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. RESEARCH · CL_228802 ·

    New methods enhance visual document question answering with adaptive retrieval and agentic restoration

    Researchers have developed new methods to improve Visual Document Question Answering (DocVQA) and Knowledge-Based Visual Question Answering (KB-VQA). ViSAR introduces an adaptive retrieval technique that dynamically sel…

  2. RESEARCH · CL_186966 ·

    New UniME-R1 framework improves multimodal retrieval with feedback-driven reasoning · 2 sources tracked

    Researchers have developed UniME-R1, a novel framework designed to enhance unified multimodal retrieval by incorporating retrieval feedback into the reasoning process. Unlike previous methods that relied solely on query…

  3. TOOL · CL_141737 ·

    TextGaze uses LVLM for gaze target estimation · arXiv cs.CV

    Researchers have introduced TextGaze, a novel architecture for gaze target estimation that utilizes a Large Vision-Language Model (LVLM) for semantic guidance. This approach aims to overcome the limitations of existing …