Linear Representation Hypothesis

Latest papers 12

All topics
CardsList
  1. LinSlot: Exploiting Linear Representation hypothesis for unsupervised attribute discovery from slot based object representation

    Oct 7, 2026Sanket Gandhi, Utkarsh Giri, Varun Subramanium +2Disentangled Representation LearningUnsupervised Learning

  2. Verifying the Linear Representation Hypothesis: How Interpretable Are Vision SAEs?

    Sep 28, 2026Teodor Chiaburu, Franz Motzkus, Frank Haußer +1Linear Representation HypothesisSparse Autoencoders

  3. The Linear Representation Hypothesis Needs a Group Action

    Sep 22, 2026Louie Hong Yao, Yuhao Li, Shengchao LiuLinear Representation HypothesisNeural Network Interpretability

  4. The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts

    Sep 21, 2026Manjiang Yu, Hongji Li, Zihan Wang +5LLM InterpretabilityLinear Representation Hypothesis

  5. How are linear representations learned? Exact solutions to the dynamics of abstraction

    Jul 9, 2026William W. Yang, Andrew M. Saxe, Peter E. LathamLinear Representation HypothesisDeep Linear Networks

  6. Closing the Loop: PID Feedback Control for Interpretable Activation Steering in Symbolic Music Generation

    Jun 17, 2026Ioannis Prokopiou, Pantelis Vikatos, Maximos Kaliakatsos-Papakostas +2Transformer InterpretabilityControllable Music Generation

  7. Representation Alignment Rests on Linear Structure

    May 22, 2026Kiril Bangachev, Guy Bresler, Yury PolyanskiyLinear Representation HypothesisCross-Modal Alignment

  8. Tensor Product Representation Probes Reveal Shared Structure Across Linear Directions

    May 11, 2026Andrew Lee, Fernanda Viégas, Martin WattenbergTransformer InterpretabilityNeural Representation Geometry

  9. The Cylindrical Representation Hypothesis for Language Model Steering

    May 3, 2026Lang Gao, Jinghui Zhang, Wei Liu +7Language Model SteeringLLM Interpretability

  10. Compositional Generalization Requires Linear, Orthogonal Representations in Vision Embedding Models

    Feb 27, 2026Arnas Uselis, Andrea Dittadi, Seong Joon OhNeural Representation GeometryVisual Representation Learning

  11. Beyond Forgetting: Representation Misdirection Elicits Controllable Side Behaviors and Capabilities

    Jan 29, 2026Tien Dang, The-Hai Nguyen, Dinh Mai Phuong +5Linear Representation HypothesisControllable Text Generation