PulseAugur
EN
LIVE 19:12:06
ENTITY depth map

depth map

PulseAugur coverage of depth map — every cluster mentioning depth map across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
6
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
6
6 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 6 TOTAL
  1. TOOL · CL_239530 ·

    New DART pretraining method enhances surgical vision models with depth data

    Researchers have developed DART, a new pretraining method for surgical vision foundation models that incorporates depth map information alongside standard RGB images. This approach, which builds upon the DINOv2 architec…

  2. TOOL · CL_216164 ·

    New VisTa3D dataset targets thin object 3D reconstruction challenges

    Researchers have introduced VisTa3D, a novel dataset and benchmark designed to improve the 3D reconstruction of thin objects. Current 3D reconstruction models struggle with thin objects due to their limited visual and p…

  3. TOOL · CL_154644 ·

    MixDiffusion framework enables multi-condition text-to-image synthesis

    Researchers have introduced MixDiffusion, a novel framework designed to enhance text-to-image generation by allowing the integration of multiple control conditions simultaneously. Unlike existing methods that are typica…

  4. RESEARCH · CL_131395 ·

    ProxyPose uses video-to-video translation for 6-DoF pose tracking

    Researchers have developed ProxyPose, a novel method for tracking the six-degree-of-freedom (6-DoF) pose of objects and surfaces from monocular video. This approach reframes the problem as a video-to-video translation t…

  5. RESEARCH · CL_99810 ·

    SpatialSV framework enhances MLLMs' 3D spatial awareness with interpretable visual supervision

    Researchers have introduced SpatialSV, a novel framework aimed at enhancing the 3D spatial awareness of multimodal large language models (MLLMs). Unlike existing methods that rely on external tools or opaque feature dis…

  6. RESEARCH · CL_36074 ·

    New frameworks tackle open-vocabulary 3D scene graph generation

    Two new research papers introduce novel frameworks for generating open-vocabulary 3D scene graphs. The first, RelWitness, addresses incomplete supervision by using visual-geometric cues to verify relations between objec…