eess.ASSep 27, 2026

Domain-Adaptive Dual-Gating Mixture of Experts for Generalizable Speech Deepfake Detection

Authors: Siqing Qin, Zhe Li, Kong Aik Lee, Man-Wai Mak

Organizations: Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University · Speech, Language, and Cognition Laboratory, The University of Hong Kong

Abstract

Recent advances in speech deepfake detection (SDD) have leveraged the Mixture of Experts (MoE) to enhance generalization capacity. However, existing gating networks often overlook the acoustic and temporal cues of deepfakes. In this work, we propose a novel domain-adaptive dual-gating MoE (DADGMoE) framework for SDD under unseen attack types and acoustic conditions. Our innovative dual-gating mechanism leverages Sinc-layer-based filters to process both low-level acoustic signals (raw waveforms) and high-level speech representations from a large self-supervised learning (SSL) model. It further incorporates domain prototypes to guide expert routing based on implicit deepfake patterns. The lightweight affine experts process the routed inputs. Experiments show that our DADGMoE significantly outperforms the baseline, achieving up to a 40.8% relative EER reduction on challenging out-of-dataset benchmarks. This framework demonstrates superior generalization capabilities and efficient design.

Figures & tables

Explore similar work

CardsList
  1. GLAD: Global-Local Adaptive Detector for Robust Speech Deepfake Detection

    Sep 28, 2026Zelin Zhao, Guanjie Huang, Danny Hin Kwok Tsang +1Audio Deepfake DetectionSpeech Synthesis

  2. DGS-MLDG: Domain Gradient Surgery Guided Meta-Learning for Domain Generalization in Speech Deepfake Detection

    Sep 27, 2026Siqing Qin, Kong Aik Lee, Youzhi Tu +2Multimodal Domain GeneralizationMeta-Learning

  3. Time-Frequency Consistency Learning for Robust Speech Deepfake Detection

    Jul 20, 2026Jun Xue, Zhuolin Yi, Yanzhen Ren +6Audio Deepfake DetectionDistortion