LLM Uncertainty Estimation

LLM: Large Language Model

Latest papers 298

All topics
CardsList
  1. Intrinsic Sequence-Likelihood Confidence in Retrieval-Dominated Extractive QA: Two Pre-Specified Negatives, and What They Do and Do Not Attribute

    Sep 17, 2026Gunwoo Lee, Changmin Sung, Sang-Hwan Gwak +3LLM EvaluationSelective Prediction

  2. Attention Dispersion as a Diagnostic Signal for Hallucination in Large Language Models

    Sep 16, 2026Shardul P. More, Tanuja S. PawarHallucination DetectionLLM Hallucination Detection

  3. Confidence Comes from Experience: Experiential Confidence Estimation from Reasoning to Agents

    Sep 15, 2026Caiqi Zhang, Xiaochen Zhu, Chengzu Li +3Confidence Estimation in Language ModelsLLM Uncertainty Estimation

  4. When Should LLMs Abstain? Chain-of-Self-Questioning for Selective Risk Control

    Sep 15, 2026Ali ŞenolSelective PredictionLLM Prompting

  5. R2VC: Modular Fact-Checking with Retrieval, Verification, and Confidence Calibration

    Sep 14, 2026Dhruv Dixit, Paritosh PandeyRetrieval-Augmented GenerationAutomated Fact-Checking

  6. ProbPlug: A Plugin Uncertainty Network for Reliable Confidence in LLM Binary Classification

    Sep 9, 2026Jianzong Wang, Chuhang Liu, Botao Zhao +6Confidence Estimation in Language ModelsLanguage Model Calibration

  7. Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning

    Sep 8, 2026Mar Gonzàlez I Català, Haitz Sáez de Ocáriz Borde, Davide Murari +3LLM Uncertainty EstimationLLM Reasoning

  8. From Answers to Interpretations: Rethinking Ambiguity-Induced Aleatoric Uncertainty Estimation in LLMs

    Sep 7, 2026Omer Nahum, Niv Nayman, Jonathan Fhima +4LLM Uncertainty EstimationAleatoric Uncertainty

  9. DualStake: Dual-Path Confidence Calibration in Deep Research Agents

    Sep 1, 2026Yinuo Xu, Yuwei Liang, Jianjie Cheng +4Deep Research AgentsLanguage Model Calibration

  10. BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs

    Aug 31, 2026Debarpan Bhattacharya, Malay Phadke, Sriram GanapathyUncertainty QuantificationLLM Uncertainty Estimation

  11. When Do Supervised UQ Ensembles Improve LLM Hallucination Detection? A Robustness Study

    Aug 25, 2026Mohit Singh Chauhan, Vipin Gyanchandani, Dylan BouchardLLM Hallucination DetectionLLM Uncertainty Estimation

  12. Credal Large Language Models for Semantic Commitment under Uncertainty

    Aug 24, 2026Shireen Kudukkil Manchingal, Sofiia Nikolenko, Fabio CuzzolinSelective PredictionConfidence Estimation in Language Models

  13. PropUQ-MAS: Propagation-Aware Uncertainty Quantification for LLM Multi-Agent Systems

    Aug 22, 2026Yaokun Liu, Yifan Liu, Daniel Yue Zhang +3Multi-Agent LLM SystemsLLM Uncertainty Estimation

  14. MAE I Trust Myself? Self-Evaluating VLA Action Generation with Markov Attention Entropy

    Aug 17, 2026Aniri, Chen Yilin, Jinhe Bi +7Uncertainty Estimation for VLA ModelsLLM Uncertainty Estimation

  15. Asymptotic Risk Calibration for Selective Question Answering

    Aug 12, 2026Shufan Lin, Sijin DongQuestion AnsweringLLM Uncertainty Estimation

  16. Attention-Path Fragility as an Uncertainty Signal in Large Language Models

    Aug 11, 2026Minsoo Kim, Sungyoung Ji, Kisung Moon +1Confidence Estimation in Language ModelsLLM Uncertainty Estimation

  17. Rethinking LLM Verification: Evidence Structure, Uncertainty, and Selective Refinement

    Aug 11, 2026Uma Ranjan, Kunal Tilaganji, Aditya Koul +9LLM GroundingLLM Answer Verification

  18. ProbGuard: Calibrated Safety Risk Estimation from LLM Output Distributions

    Aug 11, 2026Xinzhe Huang, Biwu Yao, Kedong Xiu +4LLM GuardrailsLanguage Model Calibration

  19. When Confidence Fails: Overconfidence in LLMs under Uncertainty and Missing Clinical Information

    Aug 10, 2026Maryam Tahermazandarani, Adnan Mahmood, Fahmida Islam +1Hallucination in Language ModelsLLM Uncertainty Estimation