nocaps
PulseAugur coverage of nocaps — every cluster mentioning nocaps across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Adjudicated Captioning framework boosts zero-shot image captioning performance
Researchers have developed a novel multi-agent framework called Adjudicated Captioning to improve zero-shot image captioning. This inference-time system enhances an existing captioner by adding a stronger retrieval enco…
-
New research reveals privacy risks in vision-language models
New research indicates that multi-modal vision-language models (VLMs) are susceptible to privacy attacks, specifically membership inference attacks (MIAs), which can leak sensitive training data. One study proposes a ne…
-
Researchers find single hub text exploits vulnerabilities in CLIP cross-modal encoders
Researchers have identified a vulnerability in cross-modal encoders like CLIP, which map text and images into a shared embedding space. They discovered that a single "hub text" can generate high similarity scores with n…