PulseAugur
实时 23:05:22
English(EN) Long-Horizon Audio-Visual Generation for Persistent Stories and Interactive Worlds

JoyAI-Echo-1.5系统统一长视频和交互式世界生成

研究人员推出JoyAI-Echo-1.5,一个用于生成长篇视听内容(包括持久化故事和交互式世界)的统一系统。该系统利用跨镜头记忆来保持角色身份和声音在长序列中的一致性,并结合了面向几何的相机控制,以实现世界生成中的灵活视角。JoyAI-Echo-1.5在WBench和SANA-WM-Bench基准测试中均获得最高排名,展示了其在视觉质量、文本对齐和长视域持久性方面的有效性。 AI

影响 该系统提升了AI在创建连贯、长篇叙事内容和交互式虚拟环境方面的能力。

排序理由 该集群描述了一篇关于用于视听生成的新型AI模型的研究论文。

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 3 个来源。 我们如何撰写摘要 →

JoyAI-Echo-1.5系统统一长视频和交互式世界生成

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
该集群描述了一篇关于用于视听生成的新型AI模型的研究论文。
Source corroboration
3 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
29 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+1 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

完整方法见我们的编辑标准

报道来源 [3]

  1. Hugging Face Daily Papers TIER_1 English(EN) ·

    面向持久化故事和交互式世界的长时域视听生成

    Video generation is progressing beyond isolated clips toward long-form narratives and interactive worlds, requiring models to preserve identities, follow user controls, and remain stable over extended rollouts. We present JoyAI-Echo-1.5, a unified audio-visual generation system w…

  2. Hugging Face Daily Papers TIER_1 English(EN) ·

    面向持久化故事和交互式世界的长时域视听生成

    JoyAI-Echo-1.5 unifies long-form video and interactive world generation through cross-shot memory, geometry-aware camera control, and rollout-aware training to maintain identity and coherence over extended sequences.

  3. arXiv cs.CV TIER_1 English(EN) · Nan Duan, Haoyang Huang, Weiyang Jin, Haoran Li, Yaowei Li, Yuming Li, Yijun Liu, Xin Lu, Xiaoxiao Ma, Yanwen Ma, Yaofeng Su, Yilang Sun, Haoyu Wang, Zeyue Xue, Songchun Zhang, Junhao Zhuang ·

    面向持久化故事和交互式世界的长时域视听生成

    arXiv:2608.23383v1 Announce Type: new Abstract: Video generation is progressing beyond isolated clips toward long-form narratives and interactive worlds, requiring models to preserve identities, follow user controls, and remain stable over extended rollouts. We present JoyAI-Echo…