arXiv:2512.02473v2 Announce Type: replace-cross Abstract: Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on past observations and navigation actions. However, achieving temporally and spati…
arXiv:2607.18367v1 Announce Type: new Abstract: Unlike conventional video game development, which relies on labor-intensive pipelines for asset production, animation, physics, and programming, video world models generate interactive environments from user inputs instantly. It ena…
Unlike conventional video game development, which relies on labor-intensive pipelines for asset production, animation, physics, and programming, video world models generate interactive environments from user inputs instantly. It enable us to create customized, explorable, and con…
arXiv:2509.07996v4 Announce Type: replace Abstract: World modeling has become a cornerstone in AI research, enabling agents to understand, represent, and predict the dynamic environments they inhabit. While prior work largely emphasizes generative methods for 2D image and video d…
<table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1v34mmj/alayaworld_longhorizon_and_playable_video_world/"> <img alt="AlayaWorld: Long-Horizon and Playable Video World Generation" src="https://external-preview.redd.it/N2NpYXE3cHhicGVoMYn-lh-QF10BQrOav3t…
<table> <tr><td> <a href="https://www.reddit.com/r/singularity/comments/1v34r26/alayaworld_is_a_fullstack_opensource_video_world/"> <img alt="AlayaWorld is a full-stack, open-source video world model that supports 720p, 24 FPS streaming video generation with camera control" src="…