PulseAugur
EN
LIVE 09:04:42
ENTITY MSR-VTT

MSR-VTT

PulseAugur coverage of MSR-VTT — every cluster mentioning MSR-VTT across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
6 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 8 TOTAL
  1. TOOL · CL_254945 ·

    New HyVol Module Boosts Multimodal Retrieval Accuracy

    Researchers have developed a novel training-time module called Hypergraph-Regularized Gramian Volumes (HyVol) to enhance multimodal retrieval systems. This module incorporates semantic relationships between training sam…

  2. RESEARCH · CL_233605 ·

    New datasets and models advance text-to-long-video retrieval capabilities

    Researchers have introduced MELON, a new large-scale dataset designed for multi-event text-to-long-video retrieval, addressing the limitations of existing datasets that primarily focus on short clips with single events.…

  3. TOOL · CL_196224 ·

    New WSV framework improves zero-shot video captioning with synthetic video generation

    Researchers have developed a new framework called WSV for zero-shot video captioning that addresses the cross-modal gap between text-only training and video-based inference. The method involves generating synthetic vide…

  4. TOOL · CL_180206 ·

    New PHA-Net improves text-video retrieval with prototype alignment

    Researchers have developed PHA-Net, a novel network for text-video retrieval that utilizes shared prototypes to align cross-modal representations efficiently. This approach addresses the semantic mismatch between text a…

  5. RESEARCH · CL_160961 ·

    New framework boosts video captioning accuracy without retraining models

    Researchers have developed ProCap, a novel framework designed to enhance video captioning without retraining existing large vision-language models. This method uses a lightweight scoring mechanism to identify and priori…

  6. RESEARCH · CL_160958 ·

    New framework aligns text and video distributions for improved retrieval

    Researchers have introduced the Distribution-Alignment Bridge (DAB), a novel framework for text-to-video retrieval that treats the task as a distribution alignment problem. Instead of direct matching, DAB models text an…

  7. RESEARCH · CL_63060 ·

    PEEK method efficiently selects key video frames for captioning

    Researchers have developed PEEK, an efficient method for selecting essential frames from videos for captioning. This technique distills knowledge from a larger teacher model into a smaller one, enabling it to identify t…

  8. TOOL · CL_49297 ·

    New GLCCL method enhances text-video retrieval accuracy

    Researchers have developed a new method called Global-Local Contrastive Consistency Learning (GLCCL) to improve text-video retrieval. This approach uses a parameter-free module to generate semantic features from video f…