cs.LGSep 27, 2026

Benign Overfitting for General Norms and Distributions

Authors: Daniel Barzilai, Ohad Shamir

Organizations: Weizmann Institute of Science · University of Toronto & Vector Institute

Abstract

Understanding why predictors can generalize despite interpolating noisy training data is a central puzzle in machine learning. Most work on such "benign overfitting" studies minimum-2-norm linear regression, reflecting the inductive bias of gradient descent. However, modern optimizers such as Adam and Muon use non-Euclidean update geometries, favoring solutions associated with other norms. Analyzing regression for non-Euclidean norms is substantially more difficult, with known results essentially limited to Gaussians. In this paper, we develop a method to analyze benign overfitting in linear regression for general norms and general (sub-Gaussian) distributions. As a special case, we prove that minimum-p-norm interpolation with p>1 can benignly overfit even for non-Gaussian distributions, under suitable conditions. Perhaps surprisingly, for the 1-norm, benign overfitting does not hold in general for well-behaved (but non-Gaussian) distributions, showing that existing positive 1-norm results rely crucially on Gaussianity. Our proof analyzes the geometry of the dual optimization problem, using concentration and central limit tools to show it is approximately Euclidean in many high-dimensional cases.

Figures & tables

Explore similar work

CardsList
  1. Grokking through the Lens of Minimum-Norm Interpolation

    Sep 29, 2026Gil Kur, Ileana Rugina, Clémentine Carla Juliette Dominé +1Stochastic InterpolantsInterpolation

  2. Universality of Benign Overfitting in Binary Linear Classification

    Jan 17, 2025Ichiro Hashimoto, Stanislav Volgushev, Piotr ZwiernikOverfittingEmpirical Risk Minimization

  3. How abundant are good interpolators?

    Jun 4, 2026August Y. Chen, Ahmed El AlaouiStochastic InterpolantsOverparameterization