PulseAugur
EN
LIVE 10:23:36
ENTITY MMLongBench-Doc

MMLongBench-Doc

PulseAugur coverage of MMLongBench-Doc — every cluster mentioning MMLongBench-Doc across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
4
10 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
4
10 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 15 TOTAL
  1. TOOL · CL_254854 ·

    New system enhances document Q&A with visual retrieval and evidence threading

    Researchers have developed a novel system for question answering on long-context documents, particularly those with visual elements like charts and infographics. The system, named VisRAG-Ret, utilizes a frozen Qwen2.5-V…

  2. TOOL · CL_218134 ·

    New MCite-RL framework enhances multimodal RAG with citation-enhanced reinforcement learning

    Researchers have developed MCite-RL, a new framework designed to improve the reliability of multimodal Retrieval-Augmented Generation (RAG) systems. This approach uses an agentic reinforcement learning method to enhance…

  3. TOOL · CL_206310 ·

    D2-ScaleAgent framework enhances long document understanding

    Researchers have introduced D2-ScaleAgent, a novel framework designed to enhance the understanding of long and visually rich documents. This agentic system employs a dual-dimensional scaling paradigm, dynamically adjust…

  4. TOOL · CL_205918 ·

    New Trident method enhances multimodal QA for long documents

    Researchers have developed a new method called Trident to improve multimodal question answering over long documents. Trident consists of two components: Trident-R, an LLM reranker that creates structured semantic record…

  5. RESEARCH · CL_193337 ·

    New frameworks tackle long and evolving document understanding challenges

    Researchers have developed new frameworks to tackle the challenges of understanding long and evolving documents. InSight-doc, an agentic visual perception framework, adaptively allocates visual resolution to improve acc…

  6. RESEARCH · CL_180452 ·

    New benchmarks and frameworks tackle extra-long document understanding

    Researchers have introduced two new frameworks for improving the ability of large language models to understand and answer questions from very long documents. DocTrace focuses on creating a traceable evidence graph to s…

  7. TOOL · CL_169642 ·

    New VLD-RAG framework enhances AI's ability to process long, visual documents

    Researchers have developed VLD-RAG, a novel agentic framework designed for retrieval-augmented generation over long, visually-rich documents. This system constructs a multimodal index that preserves page layout and inco…

  8. TOOL · CL_156601 ·

    New TAP-RAG framework improves multimodal QA on long documents

    Researchers have introduced TAP-RAG, a novel framework designed to enhance multimodal question answering over long documents. This system employs a Task-Aware Policy Controller (TAPC) that analyzes queries to determine …

  9. RESEARCH · CL_141429 ·

    New benchmark SynthDocBench reveals VLM failures in long-context document understanding

    Researchers have introduced SynthDocBench, a novel synthetic benchmark designed to evaluate the long-context visual document understanding capabilities of vision-language models (VLMs). Unlike existing benchmarks, Synth…

  10. RESEARCH · CL_117131 ·

    New framework learns to dynamically orchestrate AI retrievers for document reasoning

    Researchers have developed a novel framework for multimodal document reasoning agents that learns to dynamically orchestrate various retrieval methods. This failure-driven evolution approach allows a meta-agent to adapt…

  11. RESEARCH · CL_93328 ·

    MAGE-RAG framework enhances multimodal QA for long documents

    Researchers have introduced MAGE-RAG, a novel framework designed to improve multimodal question answering for long documents. This system constructs an adaptive graph of evidence, incorporating text, images, tables, and…

  12. TOOL · CL_79443 ·

    EviProp method improves long document retrieval with graph diffusion

    Researchers have developed EviProp, a novel method for retrieving relevant pages from long, visually rich documents. Unlike existing approaches that score pages independently, EviProp models documents as multimodal Chun…

  13. RESEARCH · CL_77112 ·

    New CDS method advances multimodal document question answering

    Researchers have developed a new retrieval method called Constrained Dominant Sets (CDS) for multimodal document question answering. This technique addresses limitations in current systems that struggle with long docume…

  14. RESEARCH · CL_72651 ·

    MARDoc framework enhances multimodal long document QA with structured memory

    Researchers have introduced MARDoc, a novel framework designed to improve question answering for long, multimodal documents. This system utilizes three specialized agents: an Explorer for retrieval, a Refiner for proces…

  15. TOOL · CL_46440 ·

    LLMs with Vision Capabilities Tested Against OCR for Document QA

    A benchmark compared vision-capable large language models against OCR-based pipelines for question-answering on long, image-heavy documents. The evaluation used 30 PDFs from the MMLongBench-Doc dataset, assessing the mo…