cs.LGOct 5, 2026

Loss-Invariant Projections as Passive Probes of Learned Representations

Authors: Akshay Chandrasekhar, Pavlo Melnyk

Organizations: Linköping University

Abstract

Learned feature representations in neural networks often contain structure beyond that directly used by the final task output. We study this structure using passive probes\textit{passive probes} that apply fixed, untrained, property-independent projections to representations as they evolve during training. We motivate this approach through the task of prediction on S2S^2 where equivalent vector and Hermitian parameterizations reveal an additional loss-invariant trace coordinate. This motivates a general construction in which fixed random projections serve as observers of learned features. Because the observer is loss-invariant and independent of the property being studied, changes in accessibility reflect changes in the representation relative to the fixed observer rather than adaptation of the observer itself. We show that ensembles of passive probes can directly reflect task-relevant information such as target alignment. Under our constructions, the accessibility of eventual difficulty evolves differently across tasks. It increases during training in the regression tasks of surface-normal estimation and image inpainting but remains near its initial level in image classification. Comparisons with learned linear probes further show that recoverability and passive accessibility can evolve differently during training. Together, these results show how passive probes can separately characterize changes in representation geometry and the accessibility of eventual task difficulty.

Figures & tables

Appendix figures & tables2 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Deep Minds and Shallow Probes

    May 12, 2026Su Hyeong Lee, Risi KondorNeural RepresentationsHidden States

  2. Finite Probes Suffice: Identifiability and Universality for Weight-Space Learning

    Sep 27, 2026Soutrik Sarangi, Yonatan Sverdlov, Adir Dayan +2Neural NetworkHidden States

  3. Platonic Projection Structures: Operator-Induced Observability in Representation Learning

    Jul 6, 2026Kazuo Ishii, Bishnu Prasad Gautam, Jieling Wu +1Partial ObservabilityRepresentation Space