PulseAugur
中
实时 01:27:01
English(EN) Realtime-Venus: A full-duplex interaction system with asynchronous delegation

Realtime-Venus系统实现全双工交互,采用双AI模型 · 追踪2个来源

研究人员推出Realtime-Venus,一个旨在实现更自然人机交互的新型全双工交互系统。该系统利用两个专门的9B模型:Realtime-Venus-Omni用于视听任务,Realtime-Venus-Audio用于语音交互。这些模型通过共享时间线整合了连续感知、对话控制和语音生成,允许在不中断前景对话的情况下进行异步后台推理和工具执行。Realtime-Venus-Omni在视频基准测试中表现强劲,而Realtime-Venus-Audio在音频理解和语音问答方面处于领先地位,在续接指标上优于Gemini 3.1 Live和GPT-4o等模型。 AI

影响 为交互式AI系统树立了新标准,可能影响未来的多模态和对话式代理开发。

排序理由 发布论文,详细介绍了一个新AI系统及其在基准测试中的性能。

在 Hugging Face Daily Papers 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

Realtime-Venus系统实现全双工交互,采用双AI模型 · 追踪2个来源

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
发布论文,详细介绍了一个新AI系统及其在基准测试中的性能。
Source corroboration
2 independent sources
Multiple independent publishers reporting the same story raises confidence that it's real and newsworthy.
Topics
model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
18 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [2]

  1. Hugging Face Daily Papers TIER_1 English(EN) ·

    Realtime-Venus:一个具有异步委托的全双工交互系统

    Realtime-Venus is a proactive full-duplex system with separate audio-visual and audio models that integrate continuous perception, conversational control, and native speech generation via a shared causal timeline and dual-loop runtime.

  2. arXiv cs.CV TIER_1 English(EN) · Ruixiang Zhao, Hualei Wang, Renhe Sun, Enzhi Zhou, Jincenzi Wu, Xujie Song, Kexin Shi, Zihang Liu, Pengcheng Zhu, Jiayi Zhou, Baoyue Zhang, Changhao Zhang, Zitong Wang, Jinhong Wang, Tong Niu, Jingjing Liu, Junan Lin, Haolin He, Hengshuo Chu, Yuhui Chen,… ·

    Realtime-Venus:一个支持异步委托的全双工交互系统

    arXiv:2609.13814v1 Announce Type: new Abstract: Natural interaction in digital and physical environments requires continuous perception and timely responses. Spoken dialogue relies on acoustic and linguistic cues, while video interaction also requires grounding the conversation i…