PulseAugur
EN
LIVE 06:00:46
中文(ZH) 现场直击:高昂真机成本锁死具身进化,Google 科学家用视频世界模型破局|RSS 2026

Google scientists use video world models to train robots cheaply

Researchers from Google Labs and NYU have developed a novel approach to train robots by using video generation models as a substitute for real-world interaction. This method, dubbed "World Gym," allows robots to undergo extensive, low-cost practice in a simulated environment, overcoming the high expense and risk associated with physical trial-and-error. The system uses a Diffusion Transformer architecture to generate future video frames based on prior images and control commands, enabling standardized evaluation and iterative improvement of robot policies. By leveraging this "world model," robots can learn to recover from failures and perform new tasks more effectively than through traditional supervised fine-tuning or standard simulators. AI

IMPACT This approach could significantly lower the cost and accelerate the development of embodied AI by enabling extensive virtual training.

RANK_REASON The item describes a new research methodology and system for training robots, presented at a top robotics conference (RSS). [lever_c_demoted from research: ic=1 ai=1.0]

Read on 雷峰网 (Leiphone) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Google scientists use video world models to train robots cheaply

COVERAGE [1]

  1. 雷峰网 (Leiphone) TIER_1 中文(ZH) ·

    On-site Report: High Real Robot Costs Lock Embodied Evolution, Google Scientists Break Through with Video World Models | RSS 2026

    <section style="text-align: center; margin: 0px 16px; line-height: 1.75em; display: block;"><img class="rich_pages wxw-img" src="https://static.leiphone.com/uploads/new/images/20260727/6a66c291d24cf.jpg?imageMogr2/quality/90" style="width: 100%; display: inline-block; text-align:…