PulseAugur
EN
LIVE 14:45:56
ENTITY visual question answering (VQA)

visual question answering (VQA)

PulseAugur coverage of visual question answering (VQA) — every cluster mentioning visual question answering (VQA) across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
  1. TOOL · CL_200226 ·

    New backdoor attack exploits multimodal AI models with natural language triggers

    Researchers have developed a new type of backdoor attack called Text-Guided Backdoor (TGB) that targets multimodal pretrained models. Unlike previous attacks that require specific trigger conditions, TGB utilizes natura…

  2. RESEARCH · CL_104739 ·

    New benchmarks tackle hallucination in GI endoscopy AI models

    Researchers have developed new benchmarks and datasets to address hallucination issues in vision-language models (VLMs) used for gastrointestinal endoscopy. One study introduces a benchmark using the Gut-VLM dataset to …