PulseAugur
中
实时 15:29:25
English(EN) Puppeteer: Object-Grounded Posture-Aware Co-Speech Gesture Generation

新的Puppeteer模型生成基于对象的、姿态感知的协同语音手势

研究人员开发了Puppeteer,这是一种新颖的扩散模型,旨在生成与语音协同、时间连贯、语义对齐并与周围对象相结合的手势。与以往主要关注音频-手势对齐的模型不同,Puppeteer明确纳入了姿态约束和对象几何。该模型在因果潜在空间中运行,允许进行显式的时间控制,并支持手势插值和补全等任务。为了促进评估和开发,创建了一个名为SceneGes的新合成数据集,其中包含具身协同语音手势和相应的3D对象。 AI

影响 这项研究推进了生成式AI在多模态合成方面的能力,有望改善人机交互和虚拟角色动画。

排序理由 该集群描述了一篇关于协同语音手势生成新颖模型的研究论文。[lever_c_demoted from research: ic=1 ai=1.0]

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

新的Puppeteer模型生成基于对象的、姿态感知的协同语音手势

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
该集群描述了一篇关于协同语音手势生成新颖模型的研究论文。[lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
37 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. Hugging Face Daily Papers TIER_1 English(EN) ·

    Puppeteer: 基于对象的、姿态感知的协同语音手势生成

    Generating co-speech gestures that are temporally coherent, semantically aligned with speech, and grounded with surrounding objects remains challenging. Prior speech-driven gesture models emphasize audio-gesture alignment but do not explicitly account for posture constraints or s…