PulseAugur
EN
LIVE 22:31:27
ENTITY RefCOCO+

RefCOCO+

PulseAugur coverage of RefCOCO+ — every cluster mentioning RefCOCO+ across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
3
9 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
8 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

3 day(s) with sentiment data

RECENT · PAGE 1/1 · 9 TOTAL
  1. TOOL · CL_187491 ·

    New adversarial attack targets SAM3 image segmentation models

    Researchers have developed Universal Concept Disruption (UCD), a novel adversarial attack specifically designed to target SAM3 image segmentation models. UCD learns a single image perturbation that can disrupt the model…

  2. RESEARCH · CL_183434 ·

    New Hi-Token method enhances visual grounding accuracy in AI models

    Researchers have developed Hi-Token, a novel method for generative visual grounding that improves the accuracy of bounding-box predictions by tokenizing coordinates hierarchically. This approach encodes digits for hundr…

  3. TOOL · CL_167766 ·

    StepX-Edge: On-Device UI Vision-Language Model Achieves High Accuracy

    Researchers have developed StepX-Edge, a 0.9 billion parameter vision-language model designed for on-device UI understanding. This model addresses the trade-off between accuracy and efficiency on mobile devices through …

  4. TOOL · CL_129478 ·

    New SVCR Framework Enhances Weakly Supervised Referring Expression Comprehension

    Researchers have developed a new framework called Structured Visual Compositional Representation (SVCR) to improve referring expression comprehension (REC) in weakly supervised settings. This framework explicitly models…

  5. TOOL · CL_118020 ·

    HKVLM model improves visual reasoning by separating localization from language

    Researchers have developed HKVLM, a novel approach to visual reasoning that separates localization from language generation. This model utilizes a frozen language-aligned detector and a frozen language model, connected …

  6. RESEARCH · CL_106575 ·

    CoLA framework enhances multimodal AI adaptation with dual-path LoRA

    Researchers have introduced CoLA (Cross-Modal Low-rank Adaptation), a novel framework designed to efficiently adapt foundation models for multimodal tasks. Unlike existing methods that adapt each modality in isolation, …

  7. TOOL · CL_76067 ·

    Researchers seek arXiv endorsement for Locate-SAM2 computer vision paper

    Two independent researchers are seeking an endorsement for their paper on a new computer vision system called Locate-SAM2. This system connects NVIDIA's LocateAnything-3B with Meta's SAM 2.1 through a lightweight adapte…

  8. RESEARCH · CL_55943 ·

    New Framework Enhances Semi-Supervised Segmentation with LLM Priors

    Researchers have introduced "Learning to Label" (L2L), a novel framework designed to improve semi-supervised referring expression segmentation (SS-RES) by treating pseudo-label generation as a learnable process. L2L uti…

  9. TOOL · CL_22437 ·

    Visual Para-Thinker introduces parallel reasoning to multimodal LLMs

    Researchers have introduced Visual Para-Thinker, a novel framework for parallel reasoning in multimodal large language models (MLLMs). This approach shifts from vertical scaling of reasoning depth to a parallel strategy…