PulseAugur
实时 01:32:24
English(EN) NVIDIA has released NemotronLabs VoiceChat 11B, an open full-duplex speech-to-speech model with 448ms turn-taking latency. The model processes speech continuous

NVIDIA 发布 NemotronLabs VoiceChat 11B,用于实时全双工 AI 对话

NVIDIA 推出了 NemotronLabs VoiceChat 11B,一个开源的全双工语音到语音模型,专为实时对话式 AI 设计。这个统一的模型集成了语音识别、语言理解和语音合成,实现了大约 448 毫秒的轮流延迟。一个关键特性是它能够在对话进行的同时执行工具调用,从而在不中断对话的情况下无缝集成外部功能。 AI

影响 通过降低延迟并允许在对话式 AI 应用中使用并发工具,实现更自然、更具响应性的语音交互。

排序理由 NVIDIA 发布了一个具有详细技术规格和性能指标的新开源模型。[lever_c_demoted from frontier_release: ic=2 ai=1.0]

在 Mastodon — fosstodon.org 阅读 →

AI 生成摘要 · Google Gemini · 来自 2 个来源。 我们如何撰写摘要 →

NVIDIA 发布 NemotronLabs VoiceChat 11B,用于实时全双工 AI 对话

报道来源 [2]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    NVIDIA 发布 NemotronLabs VoiceChat 11B:一个开放的全双工语音到语音模型,具有约 450 毫秒的轮流响应和实时工具调用功能

    <p>NVIDIA releases NemotronLabs VoiceChat 11B, an open full-duplex speech-to-speech model with 448 ms latency and live tool calling.</p> <p>The post <a href="https://www.marktechpost.com/2026/08/09/nvidia-releases-nemotronlabs-voicechat-11b-an-open-full-duplex-speech-to-speech-mo…

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    NVIDIA 发布 NemotronLabs VoiceChat 11B,一款开源全双工语音到语音模型,具有 448 毫秒的轮流延迟。该模型连续处理语音

    NVIDIA has released NemotronLabs VoiceChat 11B, an open full-duplex speech-to-speech model with 448ms turn-taking latency. The model processes speech continuously without chaining separate components, enabling real-time conversation with tool calling capabilities. https://www. ma…