PulseAugur
EN
LIVE 14:13:48
ENTITY VideoMAE

VideoMAE

PulseAugur coverage of VideoMAE — every cluster mentioning VideoMAE across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
6 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 10 TOTAL
  1. TOOL · CL_245690 ·

    New Arti-JEPA model adapts video models for vocal tract MRI analysis

    Researchers have developed Arti-JEPA, a new joint embedding predictive architecture designed to model real-time MRI data of the vocal tract for speech analysis. This model was trained on approximately 62 hours of unlabe…

  2. RESEARCH · CL_200267 ·

    Study finds DINOv2 effective for resource-limited self-supervised learning

    A recent study explored self-supervised learning (SSL) for image and video pretraining under resource constraints, comparing various objectives. The research found that DINOv2-style pretraining performed best with limit…

  3. TOOL · CL_193672 ·

    New Sparse Autoencoders Enhance Video Representation Interpretability

    Researchers have developed spatio-temporal sparse autoencoders (SAEs) to improve the interpretability and temporal coherence of video representations. Standard SAEs, while good at decomposing features, often sacrifice t…

  4. RESEARCH · CL_145629 ·

    AI models for heart health are spatially accurate but temporally blind

    A new research paper published on arXiv investigates the attribution methods used to explain the decisions of deep learning models in echocardiography. The study found that while these models can accurately estimate lef…

  5. TOOL · CL_121524 ·

    New Transformer framework improves distracted driver detection

    Researchers have developed a two-stage Transformer framework for accurately and efficiently localizing distracted driver behaviors in video streams. The framework combines VideoMAE for feature extraction with an Augment…

  6. RESEARCH · CL_104696 ·

    New datasets and models advance sign language recognition and translation

    Researchers have developed new methods for sign language recognition and translation. One approach uses a deep learning pipeline combining a VideoMAE video transformer for classifying sign gestures into English words an…

  7. TOOL · CL_93212 ·

    New MoFore Framework Advances Self-Supervised Video Representation Learning

    Researchers have introduced MoFore, a novel framework for self-supervised video representation learning that focuses on forecasting future latent embeddings from distant context clips. Unlike previous methods that relie…

  8. RESEARCH · CL_79496 ·

    Video foundation models show emergent intuitive physics understanding

    A new research paper investigates whether video foundation models possess an understanding of intuitive physics. The study probes frozen representations of models like V-JEPA, VideoMAE, and LTX-Video using benchmarks su…

  9. TOOL · CL_72781 ·

    New AI models detect horse eye blinks for welfare assessment

    Researchers have developed and evaluated three methods for automatically detecting and classifying horse eye blinks from video footage. These methods, including a YOLOv12 detector, an optical flow approach, and a fine-t…

  10. TOOL · CL_51554 ·

    New method steers physics reasoning in video world models

    Researchers have developed a method called physics steering to control the physical reasoning of video world models. This technique uses a linear probe's weight vector, identified as a Concept Activation Vector (CAV), w…