cs.LGFeb 2, 2026

Uncertainty Localization in LLM Reasoning via Embedding Perturbations

Authors: Qihao Wen, Jiahao Wang, Yang Nan, Pengfei He, Ravi Tandon, Han Xu

Organizations: University of Arizona · Michigan State University

Abstract

Large Language Models (LLMs) have achieved significant breakthroughs across various domains, but they can still produce unreliable or misleading outputs. For responsible LLM applications, uncertainty quantification techniques are used to estimate a model's uncertainty about its outputs, indicating the likelihood that those outputs may be problematic. For LLM reasoning tasks, it is essential to estimate uncertainty not only in the final answer but also in the intermediate reasoning process, particularly to identify where uncertainty arises. Such information may enable more fine-grained and targeted interventions during inference. In this study, we investigate which metrics can effectively localize uncertain places within an LLM reasoning trajectory. Our study reveals that uncertain intermediate continuations are more likely to occur at tokens that are highly sensitive to perturbations in the embeddings of preceding tokens. In our experiments, we show that such perturbation-based metrics achieve stronger performance in localizing uncertain intermediate steps than baseline methods, including probability-based, sampling-based, and Bayesian-based approaches. Meanwhile, our proposed metrics also enjoy good simplicity and efficiency.

Figures & tables

Appendix figures & tables13 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. U-Space: Uncovering When and Why Uncertainty Arises in Language Models

    Oct 6, 2026Tobias Braun, Nils Loose, Alexander Herzog +4Large Language Model UncertaintyLLM Reasoning Strategies

  2. The Anatomy of Uncertainty in LLMs

    Mar 26, 2026Aditya Taparia, Ransalu Senanayake, Kowshik Thopalli +1Large Language Model UncertaintyUncertainty

  3. Tracing Uncertainty in Language Model "Reasoning"

    May 8, 2026Nils Grünefeld, Bertram Højer, Philipp Mondorf +5Large Language Model UncertaintyLLM Reasoning Strategies