Full-Duplex Speech Models

Momentum

26 papers in the last four weeks, up 767% on the four weeks before. 0.3% of all new papers.

Jul 13Week of Sep 28

Latest papers 60

All topics
CardsList
  1. DuplexCadence: Exact State and Execution from a Speech Model's Declared Timelines

    Sep 28, 2026Haixiao Gao, Yimin Zheng, Linyou Xiao +1Full-Duplex Speech ModelsTime-To-First-Token

  2. SALMONN-duo: Adaptive Dual-System Coordination for Full-Duplex Voice Agents

    Sep 28, 2026Wenyi Yu, Siyin Wang, Terumi Chiba +5Full-Duplex Speech ModelsFull-Duplex Voice Agents

  3. Controlling Backchannels in Streamable Full-duplex Models

    Sep 24, 2026Maike Züfle, Peter Polák, Sefik Emre Eskimez +3Full-Duplex Speech ModelsSpeech Language Models

  4. AdaptDuplex: from static to adaptive full-duplex spoken dialogue

    Sep 24, 2026Zhiyang Zhou, Yingxin Shang, Zhou Wang +9Full-Duplex Speech ModelsDialogue

  5. A Harness for Synthesizing Diverse Naturalistic Full-Duplex Conversations

    Sep 23, 2026Matthew Sun, Vinay Kothapally, Meng Yu +4Full-Duplex Speech ModelsSynthesis

  6. Qwen-Audio-Agent Technical Report

    Sep 21, 2026Chong Deng, Yunjie Ji, Yuxiang Kong +5Full-Duplex Speech ModelsText-To-Audio

  7. MIRA: Real-Time Full-Duplex Human-Robot Interaction for Embodied Companions

    Sep 21, 2026Lijian Lin, Ye Zhu, Fan Zhang +5Human-Robot InteractionArtificial Intelligence Companions

  8. Full-Duplex Speech Models Take the Floor When Asked, Not When Needed

    Sep 17, 2026Linkai Peng, Baorian Nuchged, Kaiqi Fu +1Full-Duplex Speech ModelsUtterances

  9. SteerDuplex: Steerable Duplex Speech Dialogue Models

    Sep 14, 2026Utkarsh Tyagi, Ramaneswaran Selvakumar, Advait Gosai +13Full-Duplex Speech ModelsTurn-Taking

  10. Causal Analysis and Mitigation of Spurious Onsets in Full-Duplex Speech LLMs

    Sep 11, 2026Kento NishiSpeech Language ModelsFull-Duplex Speech Models

  11. DuplexJail: Spoken Interruption Attacks on Full-Duplex Speech Models

    Sep 8, 2026Jaechul Roh, Deepak Chandran, Amir Houmansadr +1Full-Duplex Speech ModelsAttack-Success Rate

  12. TASTE2: Text-Aligned Speech Modeling and Deployment toward Full-Duplex Voice Interaction

    Sep 8, 2026Yi-Chang Chen, Chun Wei Chen, Dien-Ruei Wu +7Full-Duplex Speech ModelsSpeech Synthesis

  13. KABURI-TTS: Phoneme-Keyed Activity-conditioned Bi-channel Utterance Rendering for Interaction

    Sep 7, 2026Ryuichiro Higashinaka, Shinnosuke Takamichi, Tetsuji OgawaFull-Duplex Speech ModelsF5-Tts

  14. DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues

    Jul 28, 2026Takyoung Kim, Kang-wook Kim, Sang Hoon Woo +3Turn-TakingFull-Duplex Speech Models

  15. Video = World + Event Stream

    Jul 16, 2026Lianghua Huang, Zhi-Fan Wu, Yupeng Shi +24StreamingMultimodal Understanding

  16. Learn2Chat: Rethinking Dyadic Talking Heads via Interaction-Modulated Monologic Priors

    Jul 11, 2026Zikai Huang, Siyue Chen, Xuemiao Xu +4Human Motion GenerationFull-Duplex Speech Models

  17. A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents

    Jul 8, 2026A. Sayyad, J. Emmons, S. Jones +2RaterFull-Duplex Voice Agents

  18. Hierarchical Acoustic-Semantic Modeling: Modality Separation and Semantic Coherence for Full-Duplex SLMs

    Jul 7, 2026Zhenyu Liu, Xuanyu Zhang, Yunxin Li +10Full-Duplex Speech ModelsAcoustic Latent Space

  19. DuplexChat: Constructing Speaker-Separated Full-Duplex Dialogue Speech at Scale for Spoken Dialogue Language Modeling

    Jul 6, 2026Wataru Nakata, Yuki Saito, Hiroshi SaruwatariFull-Duplex Speech ModelsConversational Datasets

  20. FacePlex: Toward Natural Full-Duplex Conversational Avatars

    Jun 29, 2026Habin Lim, Hah Min Lew, Jae-Ho Lee +4Precise Lip SynchronizationFull-Duplex Speech Models

  21. Wan-Streamer v0.1: End-to-end Real-time Interactive Foundation Models

    Jun 23, 2026Lianghua Huang, Zhi-Fan Wu, Wei Wang +22StreamingFull-Duplex Speech Models

  22. Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

    Jun 20, 2026Jingjing Jiang, Atsumoto Ohashi, Ryuichiro HigashinakaFull-Duplex Speech ModelsAudio-Video Generation

  23. BayLing-Duplex: Native Full-Duplex Speech Dialogue with a Single Autoregressive LLM

    Jun 12, 2026Qingkai Fang, Shoutao Guo, Yang FengFull-Duplex Speech ModelsSpeech Language Models

  24. Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

    Jun 9, 2026Atsumoto Ohashi, Neil Zeghidour, Alexandre Défossez +1Full-Duplex Speech ModelsMulti-Turn Interactions

  25. IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems

    Jun 4, 2026Tao Zhong, Jiajun Deng, Nikita Kuzmin +6Full-Duplex Speech ModelsNeural Audio

  26. DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction

    Jun 2, 2026Koki Nagano, Hongyu Liu, Seonwook Park +9Full-Duplex Speech ModelsCross-Attention

  27. Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

    May 19, 2026Pablo Riera, Pablo Brusco, Cristina Kuo +2Full-Duplex Speech ModelsTurn-Taking

  28. How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue

    May 11, 2026Hui Lu, Xueyuan Chen, Huimeng Wang +4Full-Duplex Speech ModelsDialogue

  29. DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain

    May 8, 2026Hsuvas Borkakoty, Sebastian Pohl, Cheng Wang +2Large Language Model Policy OptimizationAgentic Benchmarks

  30. PersonaKit (PK): A Plug-and-Play Platform for User Testing Diverse Roles in Full-Duplex Dialogue

    May 7, 2026Hyunbae Jeon, Jinho D. ChoiTurn-TakingFull-Duplex Speech Models

  31. Liberating LLM Capabilities in Full-Duplex Speech Models

    May 4, 2026Luoyuan Zhang, Bokai Xu, Junbo Cui +4Full-Duplex Speech ModelsSpeech Language Models

  32. MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

    Apr 30, 2026Junbo Cui, Bokai Xu, Chongyi Wang +33Omni-ModalFull-Duplex Speech Models

  33. DialogueSidon: Recovering Full-Duplex Dialogue Tracks from In-the-Wild Dialogue Audio

    Apr 10, 2026Wataru Nakata, Yuki Saito, Kazuki Yamauchi +2Full-Duplex Speech ModelsSpeech Separation

  34. RelayS2S: A Dual-Path Speculative Generation for Real-Time Dialogue

    Mar 24, 2026Long Mai, Junli LiangFull-Duplex Speech ModelsAutomatic Speech Recognition

  35. Phoenix-VAD: Streaming Semantic Endpoint Detection for Full-Duplex Speech Interaction

    Sep 24, 2025Weijie Wu, Wenhao Guan, Kaidi Wang +6Full-Duplex Speech Models