cs.SDSep 30, 2026

Synthetic Speech Attribution via Prototypical Networks

Authors: Viola Negroni, Paolo Bestagini, Stefano Tubaro

Organizations: Department of Electronics, Information and Bioengineering (DEIB), Politecnico di Milano

Abstract

Synthetic speech attribution aims to identify the generative system responsible for a speech signal, but current approaches typically rely on black-box neural networks that provide limited insight into their decisions. This work investigates prototype-based networks as an interpretable alternative, where predictions are grounded in comparisons with representative training examples. We adapt ProtoPNet to spectrogram-based speech representations and evaluate the proposed framework on the MLAAD dataset under closed-set, cross-lingual, and open-set conditions. Experiments show that prototype-based reasoning achieves competitive or improved attribution performance compared with the baseline while enabling example-based explanations. These results highlight that interpretability and performance can be jointly achieved in synthetic speech attribution through prototype-based modeling.

Figures & tables

Explore similar work

CardsList
  1. XSQ-AST: An Explainable Audio Spectrogram Transformer Framework for Localising Synthetic Speech Artifacts

    Sep 21, 2026Ben Heritage, Luca Resti, Mónica Villanueva Aylagas +3SpeakerAudio Understanding