stat.MLOct 2, 2026

Generalization Bounds for Flow-matching Generative Models for Intrinsically Low-dimensional Data

Authors: Saptarshi Chakraborty, Quentin Berthet, Peter L. Bartlett

Organizations: Department of Statistics, University of Michigan · Google DeepMind · Department of Statistics, University of California, Berkeley · Department of Electrical Engineering and Computer Sciences, UC Berkeley

Abstract

Despite the remarkable empirical success of flow-matching models, their statistical generalization guarantees remain underdeveloped. Existing analyses often impose restrictive assumptions on the estimated velocity field and yield convergence rates that fail to reflect the intrinsic low-dimensional structure common in real data, such as natural images and molecular geometries. In this work, we study the statistical generalization of flow-matching models for learning an unknown distribution PdataP_{\mathrm{data}} from finitely many samples. We derive finite-sample error bounds on the learned generative distribution, measured in the Wasserstein-pp distance, for all p≥1p\geq 1. Specifically, given nn i.i.d. samples from PdataP_{\mathrm{data}}, we show that, for every d>dp∗(Pdata)d>d_p^\ast(P_{\mathrm{data}}) and appropriately chosen network architectures and hyperparameters, the learned distribution P^FM\widehat{P}^{\mathrm{FM}} satisfies Wp(P^FM,Pdata)≲n−1/d+n−1/(2p)(log⁡(1/ξ))1/(2p) \mathbb{W}_p(\widehat{P}^{\mathrm{FM}},P_{\mathrm{data}}) \lesssim n^{-1/d}+n^{-1/(2p)}\bigl(\log(1/ξ)\bigr)^{1/(2p)} with probability at least 1−ξ1-ξ, where dp∗(Pdata)d_p^\ast(P_{\mathrm{data}}) denotes the Wasserstein-pp dimension of the target measure. Our results demonstrate that flow matching naturally adapts to the intrinsic geometry of data and mitigates the curse of dimensionality, as the convergence exponent depends on the intrinsic rather than ambient dimension. These guarantees remain meaningful in high-dimensional regimes and provide a theoretical explanation for the empirical success of flow matching on structured data distributions under substantially more relaxed assumptions than those in existing analyses.

Figures & tables

Explore similar work

CardsList
  1. Flow Matching under Noisy Latent Structure: Beyond Exact Low-Dimensional Support

    Sep 30, 2026Lifeng Hao, Shaolin JiLatent VariableRectified Linear Unit

  2. Diffusion Flow Matching: Dimension-Improved KL Bounds and Wasserstein Guarantees

    Jun 15, 2026Marta Gentiloni Silveri, Giovanni Conforti, Alain DurmusGenerative ModelsKullback-Leibler Divergence

  3. A Theory on Flow Matching with Neural Networks

    Jun 8, 2026Yihan He, Qishuo Yin, Yuan Cao +2Generative Flow NetworksVelocity Field