English(EN)Creating Impactful Autonomous Driving Datasets: A Strategic Guide from Research Gap to Benchmark
新的自动驾驶模型使用世界模型进行更安全、更鲁棒的规划 · 跟踪 2 个来源
作者PulseAugur 编辑部·[8 个来源]·
两篇新的研究论文介绍了用于端到端自动驾驶的先进世界建模技术。OWMDrive 专注于 4D 占用世界模型,用于多步 3D 占用预测,以指导基于扩散的规划,旨在实现更具前瞻性和鲁棒性的轨迹生成,尤其是在挑战性场景中。ExploreVLA 将世界建模与强化学习相结合,以实现超越专家演示的策略探索,使用未来图像生成作为密集世界建模目标和新颖性检测的内在奖励信号。
AI
arXiv cs.AI
TIER_1English(EN)·Richard Schwarzkopf, Jonas Merkert, Frank Bieder, Annika B\"atz, Alexander Blumberg, Carlos Fernandez, Felix Hauser, Fabian Immel, Christian Kinzig, Hendrik K\"onigshof, Fabian Konstantinidis, Martin Lauer, Willi Poh, Nils Rack, Kevin R\"osch, Yinzhe She…·
arXiv:2607.00710v1 Announce Type: cross Abstract: Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets contain rather than how to strategically design impactful ones. This is especially li…
Well-designed autonomous driving datasets have fundamentally shaped research progress, yet existing literature primarily describes what datasets contain rather than how to strategically design impactful ones. This is especially limiting for small and medium-sized labs and startup…
arXiv cs.CV
TIER_1English(EN)·Zhexiao Xiong, Xin Ye, Burhan Yaman, Sheng Cheng, Yiren Lu, Jingru Luo, Nathan Jacobs, Liu Ren·
arXiv:2601.04453v4 Announce Type: replace Abstract: World models have become central to autonomous driving, where accurate scene understanding and future prediction are crucial for safe control. Recent work has explored using vision-language models (VLMs) for planning, yet existi…
arXiv:2607.00399v1 Announce Type: new Abstract: End-to-end autonomous driving models often encounter performance bottlenecks, as training-time scaling leads to high computational costs and diminishing marginal returns. Existing planners typically adopt a one-shot generation parad…
Most end-to-end autonomous driving methods rely solely on instantaneous sensor observations, limiting them to reactive behavior without the anticipatory foresight human drivers employ through prior experience. We introduce geospatial visual priors, street-level visual context anc…
arXiv cs.CV
TIER_1English(EN)·Zihao Sheng, Xin Ye, Jingru Luo, Sikai Chen, Liu Ren·
arXiv:2604.02714v2 Announce Type: replace Abstract: End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies through behavior cloning on expert demonstrations. However, imitation learning in…
arXiv:2606.30421v1 Announce Type: new Abstract: Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traffic environments. However, most existing learning-based methods still make decisi…
Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traffic environments. However, most existing learning-based methods still make decisions from static representations of the current s…