PulseAugur
EN
LIVE 18:16:48

New datasets and models advance full-duplex spoken dialogue research

Researchers have introduced new datasets and models focused on full-duplex spoken dialogue systems, which enable more natural, real-time conversational interactions. The DuplexDrama dataset offers over 2,000 hours of synthesized audio data covering scenarios, expressive speech, and sound events, with a subset of bilingual dialogues to be released. SteerDuplex introduces a model and benchmark for steerable full-duplex speech, allowing control over attributes like tone and speaking rate, and demonstrating significant improvements in instruction following and interruption handling. Additionally, ConversationalVoice presents a pipeline to create training data from real conversations, generating separated, reconstructed, and expanded speech artifacts that preserve interaction dynamics. AI

IMPACT These advancements in full-duplex spoken dialogue systems could lead to more natural and responsive AI assistants and conversational agents.

RANK_REASON Multiple research papers introducing new datasets and models for spoken dialogue systems.

Read on arXiv cs.CL →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

New datasets and models advance full-duplex spoken dialogue research

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Research
Multiple research papers introducing new datasets and models for spoken dialogue systems.
Source corroboration
5 independent sources
Strong cross-source corroboration — multiple independent publishers covered this within the clustering window.
Topics
paper, model release, product
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
18 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.
Coverage growth since scoring
+2 source(s) since last score
New sources have picked up this story since our last re-score. Score will update on the next scoring pass.

Full methodology in our editorial standards.

COVERAGE [5]

  1. arXiv cs.CL TIER_1 English(EN) · Shuofeng Zhao, Hongwei Cai, Wenke Fan, Qingxiang Guo, Dawei Yang, Zhou Wang, Zhiyang Zhou, Yingxin Shang, Weixu Wang, Lin Yang, Shuran Zhou, Yang Song ·

    ECHO: A Matched-Contrast Benchmark for Context-Sensitive Turn-Taking in Full-Duplex Dialogue

    arXiv:2609.17360v1 Announce Type: new Abstract: Full-duplex spoken dialogue systems must distinguish interruptions that require yielding the floor from backchannels that permit continued speaking. Existing benchmarks typically evaluate events independently and may therefore rewar…

  2. arXiv cs.CL TIER_1 English(EN) · Ke Hu, Nourchene Ferchichi, Edresson Casanova, Ankita Pasad, Elena Rastorgueva, Chen Chen, Nithin Rao Koluguri, Piotr Zelasko, Yifan Peng, Hainan Xu, Zhehuai Chen, Boris Ginsburg ·

    Enabling Streaming User Transcription in Full-Duplex Speech-to-Speech Models

    arXiv:2609.15759v1 Announce Type: new Abstract: Full-duplex speech-to-speech (S2S) models enable natural conversational AI by allowing simultaneous listening and speaking. However, these models typically lack inherent user speech transcription, which is essential for applications…

  3. arXiv cs.CL TIER_1 English(EN) · Qingxiang Guo, Wenke Fan, Shuofeng Zhao, Dawei Yang, Zhiyang Zhou, Yingxin Shang, Hongwei Cai, Zhou Wang, Weixu Wang, Lin Yang, Shuran Zhou, Yang Song ·

    DuplexDrama: A Synthesized Dialogue Dataset with Scenarios, Full-Duplex Behaviors, Expressive Speech, and Sound Events

    arXiv:2609.12872v1 Announce Type: new Abstract: We present DuplexDrama, the first synthesized spoken dialogue dataset that simultaneously covers four dimensions: (i) complete persona and scenario settings; (ii) three full-duplex behaviors (interruption, backchannel, incomplete); …

  4. arXiv cs.CL TIER_1 English(EN) · Utkarsh Tyagi, Ramaneswaran Selvakumar, Advait Gosai, Sonal Kumar, Nikhil Barhate, Isabell Sagar, Steven Li, Miheer Bavare, Daniel Quigley, Fabiola Tapia Carrillo, Jose M Patron E, Diego Mac\'ias Guti\'errez, Paul Song, Ramani Duraiswami, Dinesh Manocha,… ·

    SteerDuplex: Steerable Duplex Speech Dialogue Models

    arXiv:2609.12623v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models support low-latency turn taking, interruption handling, and backchanneling, yet a key capability remains underexplored: steerability, the ability to reliably shift conversational behavior along a…

  5. Hugging Face Daily Papers TIER_1 English(EN) ·

    ConversationalVoice: Full-Duplex Speech Data from Real Conversations through Source-Faithful Reconstruction and Conversation-Grounded Expansion

    Full-duplex speech models require training data that preserves turn-taking, overlap, interruption, and backchannel behavior, yet these signals are entangled across speakers in noisy real-world recordings. We present Conversational Voice, a pipeline that converts real two-speaker …