PulseAugur
EN
LIVE 22:01:52
ENTITY TextVQA

TextVQA

PulseAugur coverage of TextVQA — every cluster mentioning TextVQA across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
4 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 4 TOTAL
  1. TOOL · CL_180896 ·

    New Auditing Method Assesses Visual Token Provenance in MLLMs

    A new research paper introduces a method for auditing the spatial provenance of visual tokens in multimodal large language models (MLLMs). This approach goes beyond traditional accuracy metrics to assess whether a model…

  2. TOOL · CL_179288 ·

    MAViE encoder boosts vision-language model efficiency by 80%

    Researchers have introduced MAViE, a Multi-scale Adaptive Vision Encoder designed to improve the efficiency and effectiveness of vision-language models. MAViE utilizes position-dependent gates to integrate features from…

  3. RESEARCH · CL_22506 ·

    New theory guides LLM action decisions by selecting optimal controller classes

    Researchers have introduced a "Regime Theory" to guide how large language models decide on the best action for a given input. The theory categorizes controllers into four classes, from simple fixed actions to complex pr…

  4. TOOL · CL_15761 ·

    LinMU achieves linear complexity for multimodal understanding models

    Researchers have developed LinMU, a novel Vision-Language Model (VLM) architecture that achieves linear complexity, overcoming the quadratic complexity limitations of current models. This new design utilizes an M-MATE b…