cs.LGSep 28, 2026

Weighting Schedules Govern What and When Score-Based Generative Models Learn from Multimodal Data

Authors: Jérémie Klinger, Raphaël Urfin, Giulio Biroli, Marylou Gabrié

Organizations: Laboratoire de Physique de l’École normale supérieure, ENS, Université PSL, CNRS, Sorbonne Université, Université Paris Cité, F-75005 Paris, France

Abstract

Score-based generative models generate new samples by integrating a time-dependent drift that carries Gaussian noise onto the target distribution. In practice this drift is modeled by a neural network, trained on a loss integrated over time tt with a weighting schedule w(t)w(t). Along the backward dynamics, and for multi-modal distributions, trajectories commit to modes of the target within a narrow time window, the \textit{speciation time}. In this work, focusing on high-dimensional data, we decompose the integrated loss into its single-time contributions and analyze each at fixed signal-to-noise ratio Λ(t)Λ(t): we show that Λ(t)Λ(t) sets the rate at which each feature of a multimodal target - the mode directions and their relative weights - is acquired during training. Crucially, at high Λ(t)Λ(t) all mode directions are acquired together, on a single timescale insensitive to their amplitudes, while the relative weights are not learned at all. Only near the speciation time, where Λ(t)Λ(t) becomes of order one, do all features become learnable, each on its own timescale: the weights are acquired jointly with the directions, and the directions at rates set by their relative amplitudes. For models trained on time-integrated objectives, the learning dynamics is then governed by how much of the weighting effectively sits near the speciation time, which provides insights on w(t)w(t) design choices. These results follow from an exact high-dimensional analysis of the training dynamics of unbalanced and hierarchical Gaussian mixtures. Numerical experiments on image and human genome haplotype generation recover the predicted hierarchy of learning timescales in more complex settings.

Figures & tables

Appendix figures & tables5 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. The two clocks and the innovation window: When and how generative models learn rules

    May 11, 2026Binxu Wang, Emma Lucia Byrnes Finn, Bingbin LiuGenerative ModelsClocks

  2. Non-asymptotic Convergence of Stochastic Gradient Descent in Score-based Generative Models

    Jul 6, 2026Stanislas Strasman, Sobihan Surendran, Sylvain Le CorffStochastic Gradient DescentGenerative Models

  3. Walking the Score Manifold: Continuous-time Generative Dynamics on Learned Data Manifolds

    Sep 15, 2026Jan Tauberschmidt, Brian B. Moser, Stanislav Frolov +3Generative ModelsScore-Based Diffusion Model