cs.CLSep 28, 2026

The Model Knows When to Stop: Training-Free Early Stopping for Long-Context Reading

Authors: Muath Alyobi, Mohamed Eltahir, Almoayyad Abuljdail, Riyadh Almutawa, Tanveer Hussain, Naeemullah Khan

Organizations: King Abdullah University of Science and Technology (KAUST), Thuwal, Saudi Arabia · Department of Computer Science, Edge Hill University, Ormskirk, England

Abstract

Language models often process long inputs sequentially in chunks, but continuing to read after sufficient evidence has been acquired wastes computation. Existing stopping mechanisms either learn sufficiency from internal activations or train an exit gate, while a simpler alternative asks the model whether it has read enough. We introduce Answer-Convergence Stopping (ACS), a training-free stopping rule that measures rather than asks. After each chunk, it probes the frozen model's current answer state and stops when that state is both confident and stable. The rule requires only output-side generation and token log probabilities, has no trained components, and uses one shared configuration across models and benchmarks. Because a stopping policy can save computation simply by stopping too early, we evaluate the stopping decision itself using evidence position where available. On the full LongBench-v2 with two frontier models, ACS is the only stopping policy that matches or exceeds full-reading accuracy. Furthermore, across 250 S-NIAH questions, the premature stopping rate for ACS across five models from two families ranges from 0% to 12%, compared to 8.4% to 45.6% for the verbalized gate. Taken together, ACS reveals that by properly utilizing the output signals of frozen models, we can achieve favorable behaviors like adaptive stopping without the need for additional training.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

    Jun 29, 2026Zhe Dong, Fang Qin, Manish ShahEarly StoppingLarge Reasoning Models

  2. </think> Doesn't Stop Reasoning: Analysis of Spurious CoT Termination

    Sep 3, 2026Seunghee Koh, Sungjae Choi, Minchan Kwon +2Chain-of-Thought ReasoningThink

  3. Stop When Further Reasoning Won't Help: Attention-State Adaptive Generation in Reasoning Models

    Jun 13, 2026Jiakai Li, Ke Qin, Rongzheng Wang +4Large Reasoning ModelsChain-of-Thought Reasoning