cs.LGOct 7, 2026

Eigenvalues of the Hessian in Deep Learning: The Origin of Symmetry and Its Breaking

Authors: Yossi Arjevani

Organizations: Yossi Arjevani

Abstract

Hessian spectra at trained models in deep learning exhibit a persistent pattern: eigenvalues organize into distinct clusters, including a large bulk near zero and a few isolated outliers. This paper shows that a natural account of these spectral phenomena emerges when the original setting is understood as a departure from a nearby, otherwise hidden, highly symmetric reference. Modifications, including changes to the architecture, data distribution, or parameter metric, expose a nearby reference configuration whose Hessian exhibits rich invariances-ones not accounted for by weight symmetries. There, symmetry enables a precise description of the spectra, forcing high-dimensional kernels and eigenvalues of large multiplicity. Returning to the original configuration breaks the Hessian symmetry and thereby produces the observed hierarchy of clusters and outliers. The framework is developed in some generality, with a detailed analysis of three-layer ReLU networks and applications to convolutional, graph, and transformer models, as well as to the NTK. The same mechanism is further shown to yield analogous spectral structures in layerwise Hessians and the Gauss-Newton matrix.

Figures & tables

Explore similar work

CardsList
  1. Explaining Near-Zero Hessian Eigenvalues Through Approximate Symmetries in Neural Networks

    Jul 8, 2026Marcel Kühn, Bernd RosenowReLU Neural NetworksSymmetry Breaking

  2. How the Hessian-Spectrum of Neural Networks Depends on Data

    Jul 15, 2026Jasraj Singh, Enea Monzio Compagnoni, Antonio OrvietoNeural Network OptimizationDeep Linear Networks

  3. Hessian Surgery: Class-Targeted Post-Hoc Rebalancing via Hessian Spike Perturbation

    May 8, 2026Hugo Vigna, Samuel BontempsClass-Imbalanced LearningNeural Network Optimization