PulseAugur
EN
LIVE 11:00:20
ENTITY MMStar

MMStar

PulseAugur coverage of MMStar — every cluster mentioning MMStar across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
4 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 4 TOTAL
  1. TOOL · CL_193648 ·

    New VCU-Bridge framework enhances MLLM visual reasoning hierarchy

    Researchers have introduced VCU-Bridge, a new framework designed to improve how Multimodal Large Language Models (MLLMs) understand visual information. Unlike current models that often process details and high-level con…

  2. TOOL · CL_133657 ·

    New HART technique enables LMMs to reason with high-resolution images without annotations

    Researchers have developed a new technique called HART (High-resolution Annotation-free Reasoning Technique) to improve how Large Multimodal Models (LMMs) handle high-resolution images. Current LMMs struggle with the la…

  3. RESEARCH · CL_94025 ·

    New AI Model Restores Damaged Images for Better Multimodal Understanding

    Researchers have developed Robust-U1, a novel approach to enhance the understanding of damaged images by multimodal models. Instead of solely relying on textual analysis or feature alignment, Robust-U1 generates a resto…

  4. RESEARCH · CL_04920 ·

    New CGC framework boosts multimodal LLMs for fine-grained image understanding

    Researchers have introduced Compositional Grounded Contrast (CGC), a new framework designed to enhance the fine-grained multi-image understanding capabilities of Multimodal Large Language Models (MLLMs). This approach a…