Representation Geometry in Language Models

Latest papers 221

All topics
CardsList
  1. Parts-of-Speech as Emergent Categories in SAE Latent Space

    Sep 24, 2026Alessandro Bondielli, Lucia Passaro, Serena Auriemma +1Sparse AutoencodersRepresentation Geometry in Language Models

  2. Order-Invariant Answers, Order-Sensitive Representations in Mathematical Reasoning

    Sep 23, 2026Zhixu Silvia TaoMathematical ReasoningRepresentation Geometry in Language Models

  3. The Answer-Basin Representation Hypothesis: We Are Not Probing or Steering Concepts

    Sep 21, 2026Manjiang Yu, Hongji Li, Zihan Wang +5LLM InterpretabilityLinear Representation Hypothesis

  4. Comparing Latent Concept Formation in State Space Models and Transformers via Sparse Autoencoders

    Sep 21, 2026Rithin Nagaraj, Rupa Laalasa Oruganti, Prerna Subhashchandra Kunder +1Transformer InterpretabilityLanguage Modeling

  5. Displacement Geometry Captures Platonic Shared Reality Across Models and Modalities

    Sep 21, 2026Chenming Shang, Yujin Tang, Jun Jie Ou Yang +3Cross-Modal AlignmentTransfer Learning

  6. Opinion Leader Dynamics: How Sparse Attention Shapes Token Clustering

    Sep 21, 2026Jingkun Liu, Yue SongOpinion DynamicsClustering

  7. Same Outcome, Different Readout: What Does a Steerable Valence Direction in LLMs Represent?

    Sep 19, 2026Weihan Li, Xinlei Chen, Yuhan Song +2Language Model SteeringLLM Interpretability

  8. Generalization through Lexical Abstraction in Transformer Models: The Case of Functional Words

    Sep 17, 2026Giuseppe Samo, Vivi Nastase, Paola MerloTransformer InterpretabilityWord Embeddings

  9. Large Language Models Develop Belief State Geometry In-Context

    Sep 15, 2026Daniel Balcells, Andrew Jun Lee, Chirag Rastogi +3In-Context LearningBelief Updating in LLMs

  10. Attention Mean Fields Predict Average Representation Dynamics and Reveal Context-Specific Computation

    Sep 14, 2026Micah Adler, John W. Byers, Mark CrovellaAttention Head AnalysisIn-Context Learning

  11. Disentangling Representation Evolution in Transformers through Directional Decomposition

    Sep 14, 2026Shwai He, Haichao Zhang, Shen YanNeural Representation GeometryTransformer Attention

  12. Psychosis involves a deficit of information compression in connected speech

    Sep 14, 2026Samuele Vallisa, Claudio Palominos, Rui He +6Computational PsycholinguisticsRepresentation Geometry in Language Models

  13. Implicit Personality Representations in Humans and LLMs

    Sep 14, 2026Yilin Geng, Omri Abend, Eduard Hovy +1Personality Modeling in Language ModelsRepresentation Geometry in Language Models

  14. MAxBench: A Multinomial Concept Recovery Benchmark

    Sep 14, 2026Divya Appapogu, Freya Behrens, Yonatan Belinkov +1LLM InterpretabilityRepresentation Geometry in Language Models

  15. Semantic Fibers and Cross-Gram Interference: A Calculus of Safety Drift in Overcomplete Representations

    Sep 14, 2026Mohammed Ahnouch, Lotfi ElaachackLanguage Model Safety EvaluationLLM Safety

  16. How Should Reasoning Be Organized in a Transformer's Latent Space?

    Sep 12, 2026Hongyu Gu, Chang Liu, Jingwen FuTransformer AttentionImplicit Reasoning in Language Models

  17. The information geometry of large language models is shared, learned, and controllable

    Sep 11, 2026Dario PicozziInformation GeometryLanguage Model Steering

  18. Global Divergence, Local Convergence: Representation Geometry in SSMs and Transformers

    Sep 8, 2026Amit Ben-Artzy, Roy SchwartzTransformerRepresentation Geometry in Language Models

  19. LLM Layers Immediately Correct Each Other

    Sep 7, 2026Arjun Patrawala, Jiahai Feng, Erik Jones +1Transformer InterpretabilityRepresentation Geometry in Language Models

  20. A*-Thought-V2: Efficient Latent Reasoning via Geometric Dynamics of LLM

    Sep 7, 2026Xiaoang Xu, Siyuan Liu, Shuo Wang +13Efficient Language Model ReasoningContinuous Latent Reasoning

  21. Think Wider: Mitigating Latent Rank Collapse in Implicit Chain-of-Thought Reasoning

    Sep 7, 2026Yuwen Hao, Menglin YangSpectral RegularizationContinuous Latent Reasoning

  22. Interpretable Symptom Vectors for Depression in a Large Language Model

    Sep 1, 2026Fangyi Zhu, Ajay Subramanian, Allison Constant +3Depression DetectionRepresentation Geometry in Language Models

  23. Are You Thinking What I am Thinking? : Examining Conceptual Separation in Neural Architectures

    Sep 1, 2026Jaee Ponde, Roshni Agarwal, Subhashis BanerjeeNeural Network InterpretabilityConvolutional Neural Networks

  24. Late Transformer Layers Recode Syntax Canonically: Evidence from Greek Scrambling and Cross-Layer Generalisation

    Aug 31, 2026Christos Nikolaos Zacharopoulos, Revekka Kyriakoglou, Chara Tsoukala +1Transformer InterpretabilityLanguage Modeling