Representation Geometry in Language Models

Latest papers 221

All topics
CardsList
  1. Fully Interpretable Minimal Transformers: From Geometry to Algorithm

    Oct 7, 2026Raneem Mahajne, Toviah MoldwinTransformer InterpretabilityMechanistic Interpretability

  2. When Rank Rises as LLMs Degrade

    Oct 7, 2026Zhaohui Geoffrey WangLLM Post-TrainingRepresentation Geometry in Language Models

  3. Isotropic Yet Undecodable: The Sequential Content-Sufficiency Gap in Latent-Predictive Text Representations

    Oct 6, 2026K. P. Santoso, N. Z. Fadil, F. P. Harsanti +2Non-Autoregressive Text GenerationPredictive Representation Learning

  4. A theory of platonic representations in language models

    Oct 5, 2026Darshil Doshi, Wenjie Zhou, Corinna Elena Wegner +3Multilingual Language ModelsCross-Lingual Representation Alignment

  5. A Testable Theory of Atomic Features

    Oct 5, 2026Kenny Peng, Jon Kleinberg, Nikhil GargSparse AutoencodersRepresentation Geometry in Language Models

  6. Erased, Rerouted, or Rescaled? Post-Training and the Causal Quotient of a Language Model's Belief State

    Oct 4, 2026Weihan Li, Tianshi Zheng, Junhao Wu +1Belief Updating in LLMsRL Post-Training

  7. Usage-Modulated Sentiment Representations in Large Language Models

    Oct 4, 2026Hongfei Du, Jiacheng Shi, Yanfu Zhang +2Causal Interventions in Language ModelsRepresentation Geometry in Language Models

  8. Clinical Concept Centers in LLMs

    Oct 2, 2026Aishik Nagar, Abhishek Vaidyanathan, Arun-Kumar Kaliya-Perumal +2Mechanistic InterpretabilityCausal Interventions in Language Models

  9. Beyond Linear Concepts: Discovering and Aligning Non-Linear Concept Manifolds in Large Language Models

    Oct 1, 2026Tido Specht, Elias Benedict Krey, Nils Neukirch +1LLM AlignmentLLM Interpretability

  10. Contextual trajectory and incremental contextual displacement: Towards using LLMs to understand dynamic, utterance-specific meaning construction

    Sep 30, 2026Grayson Wycliffe Storer, Julia Witte ZimmermanText EmbeddingsLanguage Modeling

  11. Concept Subspaces Compute Beyond the Logit Lens: A Weights-Only Test for Locating Representations Upstream of Readout

    Sep 30, 2026Aojie Yuan, Zhiyuan Julian Su, Haiyue Zhang +1Transformer InterpretabilityRepresentation Geometry in Language Models

  12. Targeted Retrieval, Compact Representations: How CoT Reasoning Improves Long-Context Counting

    Sep 30, 2026Liang Twist Shan, Tianyu Hu, Hao Yan +1Numerical Reasoning in Language ModelsLong-Context Retrieval

  13. S3S^3: Spectral Null-Space Swap Makes Reasoning Models Efficient

    Sep 29, 2026Hongbo Ma, Sansheng Cao, Jiajun Fan +2Efficient Language Model ReasoningModel Merging

  14. The Geometry of Inference in Transformer Residual Streams

    Sep 29, 2026Timur Mudarisov, Mikhail Burtsev, Radu StateTransformer InterpretabilityTransformer Inference

  15. Predictive Geometry of Hidden Trajectories in Transformers

    Sep 29, 2026Timur Mudarisov, Mikhail Burtsev, Tatiana Petrova +1Decoder-Only Language ModelsRepresentation Geometry in Language Models

  16. When Models Don't Manipulate Manifolds: The Geometry of a Comparison Task

    Sep 29, 2026Sai Sumedh R. Hindupur, Hadas Orgad, Thomas Fel +1Numerical Reasoning in Language ModelsMechanistic Interpretability

  17. Fisher-IRG: Fisher-Induced Local Invariant Representation Geometry across Language and Vision Models

    Sep 29, 2026Abdullah All Tanvir, Xin ZhongInformation GeometryNeural Representation Geometry

  18. Causal and Interpretable Structures in LLM Compositional Tasks

    Sep 28, 2026Gurbir Arora, Toni J. B. Liu, Jiajun Bao +2LLM InterpretabilityCausal Interventions in Language Models

  19. RoPE is Dead, Long Live RoPE: Towards Scalable Data-aware Positional Encodings

    Sep 28, 2026Jarod Lévy, Mathurin Videau, Jad Yehya +3Rotary Positional EmbeddingsLong-Context Language Modeling

  20. Learn Here, Move Less Elsewhere: Input-Conditioned Plasticity from Retained-Domain Activation Atlases

    Sep 28, 2026Jiangtao Lin, Bangyang Wei, Yihang Ding +2Continual Learning for LLMsDomain Adaptation for LLMs

  21. Rethinking Contextualization by Reinterpreting Attention Head Channels

    Sep 27, 2026Hakaze Cho, Haolin Yang, Zhun Sun +3Attention Head AnalysisRepresentation Geometry in Language Models