cs.LGOct 8, 2026

The Polytopal Neural Network

Authors: A. Emilie J. Wedenborg, Anders V. Nørskov, Teresa Dorszewski, Kristoffer Wickstrøm, Morten Mørup

Organizations: Technical University of Denmark · UiT The Arctic University of Norway

Abstract

Understanding how deep neural networks process information remains a central challenge. Existing interpretability methods often compromise structural fidelity, rely on prespecified corpora, or explain models post-hoc. We propose Polytopal Neural Networks (PNNs), a framework that extracts distinct layer-wise aspects by enforcing a polytope-based structure that is used directly in subsequent information processing. We scale our approach using learned corpus representations and an amortized simplex inference procedure and highlight how the framework also gives a direct route to vector quantized (VQ) training. In PNNs, observations are explicitly described by their alignment with layer-specific aspects. Empirical results show that imposing polytopal constraints on neural network representations preserves meaningful structures in the latent space with minimal degradation in performance, favorable compressed representations when compared to VQ representations in unsupervised learning, while also providing a performant new approach to VQ deep learning training. Our findings suggest that deep networks can enforce interpretable polytope-based representations, offering a principled path toward more transparent AI systems with minimal performance compromise.

Figures & tables

Appendix figures & tables17 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Interpretability Without Tradeoffs: Disentangling Polysemanticity At Equal Predictive Performance

    May 29, 2026Doğukan Bağcı, Bernt Schiele, Simone Schaub-Meyer +2Disentangled Representation LearningRepresentation Disentanglement

  2. Tensorization is a powerful but underexplored tool for compression and interpretability of neural networks

    May 26, 2025Safa Hamreras, Sukhbinder Singh, Román OrúsTensor NetworksNeural Network Interpretability

  3. AffineLens: Capturing the Continuous Piecewise Affine Functions of Neural Networks

    May 7, 2026Yi Wei, Xuan Qi, Furao Shen +3Neural Network InterpretabilityNeural Network Approximation Theory