RGB-D Visual Simultaneous Localization and Mapping (SLAM) Application
PulseAugur coverage of RGB-D Visual Simultaneous Localization and Mapping (SLAM) Application — every cluster mentioning RGB-D Visual Simultaneous Localization and Mapping (SLAM) Application across labs, papers, and developer communities, ranked by signal.
6 day(s) with sentiment data
-
Robotics research introduces outcome-based representation learning
Researchers have developed a novel method for learning manipulation-sufficient representations in robotics, focusing on action outcomes rather than dense geometric states. This approach utilizes an action-conditioned ou…
-
New E-RGB-D System Achieves Real-Time Color and Depth Perception at High Speeds
Researchers have developed E-RGB-D, a novel system that combines event-based cameras with a digital light processing projector to achieve real-time, high-speed RGB-D perception. This integration addresses the limitation…
-
Hydra method achieves marker-free hand-eye calibration with improved accuracy
Researchers have developed a new marker-free hand-eye calibration method called Hydra, which utilizes RGB-D imaging and a novel iterative closest point (ICP) algorithm. This approach formulates a robust point-to-plane o…
-
VOIM system generates 3D instance maps without training
Researchers have introduced VOIM, a novel training-free system for creating 3D instance maps from RGB-D or monocular RGB data. Unlike existing online systems that commit to labels early, VOIM defers these decisions unti…
-
DINOcular framework learns visuospatial representations from RGB-D data
Researchers have introduced DINOcular, a novel self-supervised framework designed to learn visuospatial representations from RGB-D (color and depth) data. This approach integrates geometric priors derived from depth inf…
-
Stream3Dv2 framework enhances zero-shot 3D scene understanding with geometric-semantic fusion
Researchers have developed Stream3Dv2, a novel framework for robust, zero-shot 3D scene understanding using streaming RGB-D inputs. This training-free approach addresses limitations in handling sequential data and noise…
-
New S3AM framework enhances multi-modal salient object detection
Researchers have developed S$^3$AM, a novel single-stream framework for multi-modal salient object detection. This approach integrates a reliability-calibrated frequency adapter with the Segment Anything Model (SAM) bac…
-
New system InViStream tackles privacy in real-time volumetric video
Researchers have developed InViStream, a novel system designed to preserve privacy in real-time volumetric video streaming. Unlike traditional video, volumetric video captures 3D scenes from multiple angles, making priv…
-
New GESTO memory system enables robots to reason about human activities
Researchers have introduced GESTO, a novel spatio-temporal memory system designed for robots operating in dynamic human environments. GESTO integrates a persistent 4D scene graph with a hierarchical structure of atomic …
-
New Energy-Structured World Model Enhances Physically Consistent Motion Planning
Researchers have developed a novel Energy-Structured Latent World Model (ELWM) to improve physically consistent motion planning in embodied AI. This model explicitly encodes energy and momentum within its latent state, …
-
GOPI framework improves 3D furniture insertion from single-view images
Researchers have developed GOPI, a novel framework for generating plausible 3D poses of furniture for insertion into indoor scenes from single-view RGB-D images. The method addresses the inherent underdetermination of s…
-
New FOCUS framework boosts MLLM salient object detection capabilities
A new research paper proposes FOCUS, a novel framework designed to enhance salient object detection (SOD) capabilities in multimodal large language models (MLLMs). The paper introduces SaliLLM, a diagnostic benchmark th…
-
New research tackles affordance segmentation for wearable robots · 2 sources tracked
Two research papers explore methods for improving visual affordance segmentation on embedded devices, particularly for wearable robots. The first paper focuses on enhancing the decoder module of lightweight neural netwo…
-
New perception pipeline aims to automate mining rock-breakers
Researchers have developed a real-time RGB-D perception pipeline to automate the operation of hydraulic impact hammers, commonly known as rock-breakers, in mining. This system integrates image-based instance segmentatio…
-
Seg2Grasp pipeline enhances robotic bin picking with modular approach
Researchers have developed Seg2Grasp, a novel modular pipeline for robust suction grasping in bin picking tasks. This system employs a three-step process: segmentation using a Transformer-based model to create object ma…
-
AlayaWorld advances interactive video world modeling with 720p generation
Researchers have introduced AlayaWorld, an interactive video world model capable of generating 24-fps video at 540p and 720p resolutions. This model utilizes a 15B video diffusion transformer and incorporates several me…
-
DA-Fusion Transformer enhances unseen object segmentation for logistics
Researchers have developed DA-Fusion, a novel Transformer model that uses deformable attention to fuse RGB and depth data for improved unseen object instance segmentation. This advancement is particularly beneficial for…
-
Depth data boosts surgical vision foundation models, study finds
A new study explored the impact of incorporating depth information into vision foundation models for surgical applications. The research found that models pre-trained with RGB-D data, such as MultiMAE, significantly out…
-
New TACTIC controller enhances whole-arm manipulation with tactile and vision data
Researchers have developed TACTIC, a new controller designed for whole-arm manipulation tasks that involve complex contact dynamics. This system integrates RGB-D vision, distributed tactile sensing, and a proximity repr…
-
GenVid2Robot framework translates generated video motion into executable robot trajectories
Researchers have developed GenVid2Robot, a framework that translates generated video motion into executable robot manipulation trajectories. This system addresses the limitations of using generated videos directly for r…