cs.LGSep 23, 2026

Thinking Leakage: A Causal Audit of NoThink Post-Training in Hybrid Reasoning Models

Authors: Zehao Liu, Vasant G. Honavar

Organizations: College of Information Sciences and Technology Pennsylvania State University

Abstract

Post-training hybrid reasoning models in NoThink mode has attracted growing interest as a way to improve performance while keeping inference fast. However, these gains may draw on thinking behavior already accessible through the base model's Think mode. We formulate this thinking leakage in a causal mediation framework and audit its contribution using bidirectional interventions along a simple base-derived activation direction. Across three models and three post-training methods on competition math benchmarks, we find that leakage is real, causal, and substantial: behavioral and representational analyses reveal shifts toward Think, steering the base model along this direction reproduces most of the post-training accuracy gain, and counter-steering a checkpoint removes a substantial share of what it gains. Across nine aligned checkpoints with positive NoThink gains, the resulting leakage ratio ranges from 42% to 79%. These interventions support a substantial causal contribution of thinking leakage. Our findings show that a post-training method's apparent advantage can therefore reflect greater drift toward Think, obscuring whether it improves capability within NoThink or more effectively re-invokes existing Think behavior.

Figures & tables

Appendix figures & tables21 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Separation

    Apr 29, 2026Shouren Wang, Wang Yang, Chuang Ma +7ThoughtsLock-In

  2. When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning

    May 21, 2025Rongzhi Zhu, Yi Liu, Jiancheng Wang +6Large Reasoning ModelsBehavioral Divergence

  3. When2Think: Learning When and How Much to Reason

    Sep 17, 2026Jaejun Shim, HyunJin Kim, Young Jin Kim +1Large Reasoning ModelsVerifiable Rewards