3D visual grounding
PulseAugur coverage of 3D visual grounding — every cluster mentioning 3D visual grounding across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
New dataset and framework enhance 3D visual grounding for autonomous driving
Researchers have introduced Talk2Sensors, a novel dataset and framework for 3D visual grounding in autonomous driving that leverages multiple sensor modalities. The dataset includes over 8,000 language instructions and …
-
New VLM-IE3D framework boosts 3D spatial understanding in vision-language models
Researchers have introduced VLM-IE3D, a novel framework designed to enhance the 3D spatial awareness of vision-language models (VLMs). This framework integrates both implicit and explicit 3D geometries derived from RGB …
-
OpenGround framework enhances 3D visual grounding with planning and online perception
Researchers have introduced OpenGround, a novel framework designed for open-world 3D visual grounding. This system addresses limitations in current methods by integrating Task-Chain Planning to break down complex querie…
-
PruneGround framework enhances 3D visual grounding with spatial pruning
Researchers have introduced PruneGround, a novel framework designed to improve 3D Visual Grounding by focusing on language-relevant regions within 3D scenes. This approach utilizes Language-Guided Spatial Pruning (LGSP)…
-
New RL framework boosts 3D video scene understanding
Researchers have introduced 3D-RFT, a novel framework that applies Reinforcement Learning with Verifiable Rewards (RLVR) to video-based 3D scene understanding. Unlike traditional Supervised Fine-Tuning (SFT) methods tha…
-
AgentGrounder enables zero-shot 3D visual grounding on point clouds
Researchers have introduced AgentGrounder, a novel framework for zero-shot 3D visual grounding that operates directly on colored point clouds. This approach bypasses the need for task-specific 3D training by employing a…