cs.LGMay 31, 2026

A Fiber Criterion for Representation Identifiability in Supervised Learning

Authors: Vasileios Sevetlidis

Organizations: Athena Research Center, Kimmeria Campus, Xanthi, Greece · Democritus University of Thrace, Vas. Sofias Campus, Xanthi, Greece · International Hellenic University, Serres, Greece

Abstract

Supervised learning evaluates predictors through their input-output behavior. When a predictor is implemented as a composition f=chf=c\circ h, supervised evidence constrains the composite map ff but need not determine the representation-head factorization (h,c)(h,c). This paper formalizes the resulting representation-level identifiability problem: for a class of admissible representation-head pairs, a representation property is identifiable from the induced predictor exactly when it is constant on the fibers of the projection (h,c)ch(h,c)\mapsto c\circ h, equivalently when it descends to a well-defined property of the predictor. Predictor-preserving augmentation gives a canonical obstruction: auxiliary information can be appended to a representation while the head ignores it, leaving the predictor unchanged but altering properties such as minimality, compression, invariance, equivariance, nuisance information, or semantic accessibility. This construction separates representation identifiability from optimization and finite-sample estimation. Finite-sample diagnostics illustrate, rather than prove, the criterion: exact algebraic witnesses hold the predictor fixed while changing representation diagnostics, and matched-performance Waterbirds models show that different constraints can select different representations at similar supervised performance. The results clarify that representation-level claims require assumptions, objectives, measurements, or inductive biases beyond supervised predictive behavior alone.

Explore similar work

Jun 2, 2026cs.LG

Bayes-Sufficient Representations in Supervised Learning

Representation learning is often described as preserving the information in an input that is relevant for prediction. This work asks what relevance means for a fixed supervised decision problem. A representation is defined to be Bayes-sufficient for a joint distribution and loss if some prediction head can use it to implement a Bayes-optimal action rule. This makes the target information loss-dependent. In the almost-surely unique Bayes-action case, the relevant object is a Bayes quotient, which identifies inputs that require the same Bayes-optimal action. A representation is sufficient when it refines this quotient, and Bayes-minimal when it is informationally equivalent to it. The framework connects naturally to property elicitation: zero-one loss requires the Bayes class, squared loss the conditional mean, Brier loss the conditional probability in binary prediction, and log loss or strictly proper scoring rules the predictive distribution. Controlled finite experiments, learned neural bottleneck experiments, and a real-data iNaturalist taxonomic refinement experiment illustrate the distinction between sufficiency, minimality, and retained non-required information. For a fixed supervised problem, the distribution and the loss determine the Bayes action, the Bayes action determines the quotient, and the quotient determines the minimal information required for Bayes-optimal prediction.
Vasileios Sevetlidis
May 12, 2026cs.LG

From Generalist to Specialist Representation

Given a generalist model, learning a task-relevant specialist representation is fundamental for downstream applications. Identifiability, the asymptotic guarantee of recovering the ground-truth representation, is critical because it sets the ultimate limit of any model, even with infinite data and computation. We study this problem in a completely nonparametric setting, without relying on interventions, parametric forms, or structural constraints. We first prove that the structure between time steps and tasks is identifiable in a fully unsupervised manner, even when sequences lack strict temporal dependence and may exhibit disconnections, and task assignments can follow arbitrarily complex and interleaving structures. We then prove that, within each time step, the task-relevant latent representation can be disentangled from the irrelevant part under a simple sparsity regularization, without any additional information or parametric constraints. Together, these results establish a hierarchical foundation: task structure is identifiable across time steps, and task-relevant latent representations are identifiable within each step. To our knowledge, each result provides a first general nonparametric identifiability guarantee, and together they mark a step toward provably moving from generalist to specialist models.
Yujia Zheng, Fan Feng, Yuke Li +3
Jun 11, 2026cs.LG

Detecting Explanatory Insufficiency in Learned Representations: A Framework for Representational Vigilance

Learned representations are central to modern machine learning and are commonly evaluated through predictive performance, robustness, uncertainty estimation, and generalization. However, a representation may remain operationally successful while failing to organize persistent residual structures that conventional metrics do not fully capture. This article introduces VER, the Vigilant Evaluator of Representations, a conceptual framework for monitoring representational adequacy. VER does not propose a new learning algorithm, loss function, or model architecture. It defines a diagnostic process for identifying residual structures and assessing whether they may indicate explanatory insufficiency rather than ordinary error, uncertainty, noise, data limitation, or distribution shift. The framework comprises five operations: representation identification, explanatory-domain delimitation, residual-structure detection, explanatory-resistance evaluation, and vigilance signaling. VER distinguishes stable adequacy, a vigilance condition, and a representational alert. It is intended to complement performance evaluation, uncertainty estimation, out-of-distribution detection, robustness analysis, and explainable AI by making representational adequacy an explicit object of inquiry. The article also outlines a path toward empirical evaluation through benchmarks designed to detect representational inadequacy when predictive performance remains satisfactory. VER is conceptual and methodological; it does not prove that a representation is inadequate, select a replacement representation, or provide an operational implementation.
Jacques Raynal, Pierre Slangen, Elsa Raynal +1