ENTITY
visual question answering (VQA)
visual question answering (VQA)
PulseAugur coverage of visual question answering (VQA) — every cluster mentioning visual question answering (VQA) across labs, papers, and developer communities, ranked by signal.
Total · 30d
0
1 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
1 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 2 TOTAL
-
New backdoor attack exploits multimodal AI models with natural language triggers
Researchers have developed a new type of backdoor attack called Text-Guided Backdoor (TGB) that targets multimodal pretrained models. Unlike previous attacks that require specific trigger conditions, TGB utilizes natura…
-
New benchmarks tackle hallucination in GI endoscopy AI models
Researchers have developed new benchmarks and datasets to address hallucination issues in vision-language models (VLMs) used for gastrointestinal endoscopy. One study introduces a benchmark using the Gut-VLM dataset to …