cs.LGSep 30, 2026

Whitening Improves Robustness to Spurious Correlations in Linear Probes

Authors: Floris Holstege, Bram Wouters, Noud van Giersbergen, Cees Diks

Organizations: University of Amsterdam, Department of Quantitative Economics · Tinbergen Institute

Abstract

Deep neural networks tend to rely on simple features that may be spurious and thus fail to generalize. We study this problem in the setting of linear probes, where a (generalized) linear model is fitted on the representations of a (pretrained) model. We use the connection of these models to the max-margin classifier, and show they favor directions associated with large eigenvalues of the covariance matrix. Whitening removes this preference by equalizing the eigenvalues of the covariance matrix. This observation motivates whitening as a preprocessing step that can reduce reliance on spurious correlations without requiring prior knowledge of their presence or labeled data. We examine the effect of whitening on a synthetic data-generating process and standard spurious correlation benchmarks, and find that it improves robustness. We also find that whitening can improve robustness when added to existing approaches.

Figures & tables

Appendix figures & tables7 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Spectral Overfitting in Noisy Linear Probing of Pretrained Representations

    Mar 2, 2026Zice Wang, Zhenyu ZhangNoisy LabelsOverfitting

  2. Comparing Linear Probes with Mahalanobis Cosine Similarity

    Jun 17, 2026Zhuofan Josh Ying, Peter Hase, Nikolaus KriegeskorteCosine SimilarityInterpretability

  3. Mitigating Spurious Correlations with Memorization-Guided Dataset De-Biasing

    Jun 1, 2026Arda Fazla, Abolfazl HashemiSpurious CorrelationsSelection Bias