cs.SDOct 5, 2026

Revisiting Label-Free Speaker Embedding Enhancement with vMF Profile Likelihood

Authors: Seunghwan Kim, Jinyong Kim, Sooyoung Yang, Youngjin Ko, Myungjoo Kang

Organizations: Interdisciplinary Program in Artificial Intelligence, Seoul National University, South Korea · Department of Mathematical Sciences, Seoul National University, South Korea · Research Institute of Mathematics, South Korea

Abstract

Embedding enhancement improves speaker verification under acoustic mismatch without modifying a frozen backbone. Recent work has established a practical label-free setting for this task, but often adopts increasingly structured formulations. Here, the clean target is directly observed during training, making enhancement a matching problem on the unit hypersphere. We model the clean target with a von Mises--Fisher (vMF) likelihood and profile out a sample-wise concentration parameter, yielding a simple closed-form objective with adaptive weighting. Across VoxCeleb1, VoxSRC23, CN-Celeb, VOiCES, and VC-Mix, the proposed method largely preserves the baseline and gives clearer gains on challenging mismatch sets. It also remains stable under a broad single-view recipe, where a recent diffusion baseline becomes less reliable in controlled comparisons. These results suggest that effective label-free embedding enhancement in this setting does not require a highly structured formulation.

Figures & tables

Explore similar work

CardsList
  1. LISE : Listenable Interpretable Speaker Embeddings

    Jun 19, 2026Xiaoliang Wu, Chongxin Gan, Ke Liu +2Automatic Speaker VerificationSpeaker

  2. ReDimNet2+: Multi-Corpus Data Scaling for Robust Speaker Verification

    Sep 29, 2026Kirill Borodin, Vasilii Kudryavtsev, Maxim Maslov +1Automatic Speaker VerificationSpeaker