研究人员推出 DiVA,这是一种新颖的数字生命模拟器,专为长期、开放式交互体验而设计。DiVA 利用多模态大型语言模型 (MLLM) 作为路由器,并结合专门的堆叠视频管道,实现无缝的多轮交互。该系统采用锚定视频续接 (AVC) 模块来管理视频片段之间的过渡,确保连续性并减少视觉退化,从而实现复杂姿势变化以及高保真度的身份和连贯性。 AI
影响 该系统可以为培训、娱乐和研究提供更逼真、更具吸引力的虚拟环境。
排序理由 该条目是一篇详细介绍新模拟系统的学术论文。[lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- Anchored Video Continuation
- arXiv
- CatalyzeX Code Finder for Papers
- Connected Papers
- CORE Recommender
- DagsHub
- DiVA
- Gotit.pub
- Hugging Face
- Influence Flower
- Litmaps
- Multimodal Large Language Model
- ScienceCast
- scite Smart Citations
AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →