cs.AISep 30, 2026

When Reasoning Goes Astray: Attention Dynamics of Uncontrolled Reasoning

Authors: Yuanhe Zhang, Ziwei Wang, Jie Ren, Haoran Gao, Zhenhong Zhou, Fanyu Meng, Cong Wu, Li Sun, +1 more

Organizations: Beijing University of Posts and Telecommunications · Wuhan University · Nanyang Technological University · Chongqing University of Posts and Telecommunications

Abstract

Large reasoning models (LRMs) improve performance on complex tasks through extended reasoning, yet the same process can degenerate into redundant verification and persistent generation loops. Such uncontrolled reasoning increases inference cost and creates risks of resource exhaustion and service degradation. However, existing mitigations largely truncate long outputs or react to surface repetition, and thus fail to distinguish normal thinking from uncontrolled reasoning or explain how benign reasoning degenerates into harmful behavior. In this paper, we operationalize LRM generation as four states and further introduce Reasoning-state Analysis via Dynamic Attention Responses (RADAR), which identifies the current reasoning state in real time and characterizes how effective reflection can develop into uncontrolled generation. Guided by RADAR's analysis, we further realign abnormal attention distributions toward patterns observed in normal requests and examine how this correction affects excessive reflection and persistent looping. Temporal analyses show that uncontrolled reasoning is characterized by attention distributions that deviate from normal generation, with abnormal trends becoming detectable before repetition begins. Correcting these deviations through Attention Realignment consistently reduces looping while largely preserving benign performance. Together, RADAR provide a mechanistic account of how reasoning becomes uncontrolled, offering actionable guidance for identifying critical failure stages and designing targeted runtime interventions.

Figures & tables

Appendix figures & tables17 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. From "Aha Moments" to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

    Aug 6, 2025Rui Ha, Rui Pu, Chaozhuo Li +2Large Reasoning ModelsMetacognition

  2. When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning

    May 21, 2025Rongzhi Zhu, Yi Liu, Jiancheng Wang +6Large Reasoning ModelsBehavioral Divergence

  3. Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models

    Jun 1, 2026Simone Caldarella, Davide Talon, Rahaf Aljundi +2Large Reasoning ModelsReasoning Traces