cs.CLSep 24, 2026

Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs

Authors: Pavel Tikhonov, Anton Korznikov, Matvey Mikhalchuk, Nikita Dragunov, Temurbek Rahmatullaev, Polina Druzhinina, Anton Razzhigaev, Ivan Oseledets, +1 more

Abstract

While Large Language Models (LLMs) rely on highly non-linear components, in this work we demonstrate that they exhibit fundamental linearity: when inputs from distinct text streams are linearly combined, the model outputs a superposition of the individual next-token distributions. We term this the \textit{Superposition Linearity Hypothesis}. We provide evidence that superposition is an intrinsic property of the Transformer architecture rather than an emergent consequence of training; in fact, we observe that it tends to diminish as pretraining progresses. However, we demonstrate that linearity can be substantially restored through lightweight fine-tuning, significantly reducing the divergence between the predicted next-token distribution and the average of the individual next-token distributions. Finally, we introduce a guided decoding procedure that disentangles superposed outputs, enabling the simultaneous generation of two coherent continuations from a single forward pass.

Figures & tables

Appendix figures & tables9 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. SuperThoughts: Reasoning Tokens in Superposition

    Jun 11, 2026Zheyang Xiong, Shivam Garg, Max Yu +4Chain-of-Thought ReasoningThink

  2. Line-Coupled Language Model

    Sep 7, 2026Shiyuan Li, Shaorong Zhang, Zhaorui Yang +3Autoregressive Language ModelsLanguage Modeling

  3. Functional Subspace, where language models can use vector algebra to solve problems

    Feb 2, 2026Jung H. Lee, Sujith VijayanSubspaceIn-Context Learning