cs.LGOct 1, 2026

Rethinking the Information Bottleneck: Structured Decomposition under Label-Induced Partitions

Authors: Jingyao Zhang, Yuxuan Li, Lu Han, Ali Anaissi, Nguyen H. Tran

Organizations: School of Computer Science, The University of Sydney Sydney, NSW 2006, Australia

Abstract

Standard information bottleneck (IB) regularization constrains representations via a single scalar I(Z;X), implicitlytreating all information as homogeneous. However, a single global compression control couples label-relevant structurewith residual within-condition variation, rather than regulating their allocation independently, allowing nuisanceinformation to persist in learned representations. For example, in medical imaging applications, residual variation oftenstems from acquisition conditions, background factors, or subject-specific appearance. This issue becomes particularlypronounced in data-limited settings, where models tend to overfit such variation, hindering generalization. While existingregularization methods can stabilize training, control capacity, or shape representation geometry, they do not explicitlyseparate nuisance-like variation from task-supporting structure. To address this limitation, we revisit IB from a structuredperspective based on a label-induced partition, where condition-level structure and within-condition information playdistinct roles. This leads to a dual-bottleneck formulation: a standard KL term controls global information capacity, while aconditional KL term targets within-condition information. We show that the conditional KL admits an exact decompositioninto a within-condition information term and a prior-mismatch term, explaining its alignment with the design objective.With a simplex-structured conditional prior, the method provides controllable latent geometry and integrates seamlesslyinto existing pipelines. Experiments on classification and segmentation show the clearest gains in low-data classificationand consistent improvements across dense prediction benchmarks.

Figures & tables

Appendix figures & tables22 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Geometry as a Missing Axis of Representation Quality: The Variational Geometric Information Bottleneck under Data Scarcity

    Nov 4, 2025Ronald KatendeDiscriminative Congruence TransformInformation Bottleneck

  2. Condensing Large-Scale Datasets Directly with Minimal Information Loss

    Jul 1, 2026Xinyi Shang, Peng Sun, Bei Shi +2Diffusion-Based Dataset DistillationDataset Distillation

  3. Geometric and Information Compression of Representations in Deep Learning

    Jun 19, 2026Linara Adilova, Henning Petzka, Asja Fischer +1Information BottleneckMutual Information