cs.LGOct 5, 2026

Sample-Optimal Estimation of the Fréchet Inception Distance

Authors: Ziyun Chen, Jerry Li, Kevin Tian, Yusong Zhu

Organizations: University of Washington · University of Texas at Austin

Abstract

The Fréchet Inception Distance (FID) is widely used to evaluate generative models, but its empirical plug-in estimator suffers from finite-sample bias [BSAG18, CF20]. We study the sample complexity nn of estimating FID to error εε between dd-dimensional Gaussians with bounded mean distance and covariances, when one distribution is known. Our contributions are threefold. (1) We establish tight finite-sample Θ(d2n)Θ(\frac{d^2}{n}) bias and Θ(dn+d2n2)Θ(\frac{d}{n} + \frac {d^2} {n^2}) variance bounds for the empirical plug-in estimator, establishing a ≳d2\gtrsim d^2 sample complexity. (2) To debias the empirical plug-in estimator, we generalize the FID∞{\rm FID}_\infty estimator of [CF20] to extrapolation methods of arbitrary order kk. We further prove tight bias and variance bounds of Θ(dk+2nk+1)Θ(\frac{d^{k + 2}}{n^{k + 1}}) and Θ(dn+d2n2)Θ(\frac d n + \frac{d^2}{n^2}) for any order-kk extrapolation under our framework. (3) We introduce Relative Taylor Debiasing (RTD), a new, computationally efficient FID estimation algorithm using debiasing techniques inspired by U-statistics. We show that RTD achieves an O(dε2)O(\frac d {ε^2}) sample complexity, and prove that this is optimal. We provide a complementary empirical evaluation of our new estimators. Our experiments on synthetic Gaussians validate the predicted residual bias and support the tightness of our bounds. On ImageNet with Inception embeddings, RTD achieves the lowest mean estimation error at the standard 50K sample budget, while our second-order variance-aware extrapolation estimator (VALE2_2) uses only 10K samples to achieve accuracy comparable to FID∞_\infty at 50K samples.

Figures & tables

Appendix figures & tables17 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. MIND: Monge Inception Distance for Generative Models Evaluation

    May 7, 2026Quentin Berthet, Yu-Han Wu, Clement Crepy +3Image Generation EvaluationSliced Wasserstein Distance

  2. The FID Lottery: Quantifying Hidden Randomness in Generative-Model Evaluation

    Jun 18, 2026Nicolas Dufour, Alexei A. Efros, Patrick PérezML ReproducibilityImage Generation Evaluation