cs.CLMay 19, 2026

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

Authors: Pablo RieraPablo BruscoCristina KuoMarcelo SancinettiS. R. K. Branavan

Organizations: ASAPP Inc., USA

Abstract

Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-based systems. Inspired by neural coupling in human communication, we study how such models coordinate their internal representations during interaction. We simulate full-duplex dialogues between two instances of the pretrained \textit{Moshi} model under controlled conditions, manipulating channel noise and decoding bias. Synchronization is measured using Centered Kernel Alignment (CKA) across temporal lags, while anticipatory turn-taking cues are probed from delayed internal activations using causal LSTM models, from both speaker and listener perspectives. We find strong representational synchronization under no noise conditions, peaking near zero lag and degrading with noise, and we show that internal states encode anticipatory information that supports turn-taking prediction ahead of time.

Explore similar work

CardsList
  1. Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models

    Jun 9, 2026Atsumoto Ohashi, Neil Zeghidour, Alexandre Défossez +1Full-Duplex Speech Models

  2. DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues

    Jul 28, 2026Takyoung Kim, Kang-wook Kim, Sang Hoon Woo +3Turn-TakingDialogue