Language Model Error Detection

Latest papers 41

All topics
CardsList
  1. Beyond Imitation: A Framework and Benchmark for LLM-Assisted Peer Review

    Oct 8, 2026Rachel S. Y. Teo, Yutaro Yamada, Shashank Kotyan +2Automated Peer ReviewLLM Evaluation

  2. Understanding Errors in LLM-Based Question Answering over Imperfect Tables

    Oct 3, 2026Baowen Zhang, Wei Fan, Ruman Wang +1Table QALanguage Model Error Detection

  3. Locating Answer-Correctness Signals in Frozen Large Language Models

    Sep 29, 2026Yuansen Liu, Yixuan Tang, Anthony Kum Hoe TungLLM Answer VerificationLanguage Model Error Detection

  4. SemOPT: Fixing Semantic Errors in LLM-based Optimization Modeling via Reward-Guided Search

    Sep 29, 2026Zetong Zhou, Wentao Zhang, Jingyuan Wang +3Language Model Error DetectionLarge Language Model-Guided Optimization

  5. The Error You See Is Not the Error You Made: Progression-aware Reasoning Origin for Reasoning Error Localization

    Sep 27, 2026Yiguo Wang, Ziyuan Yang, Yi Zou +3Language Model Error DetectionReasoning Verification

  6. Pinocchio: Fast Uncertainty Estimates for Black-Box Language Models

    Sep 21, 2026Kevin David Hayes, Arka Pal, Haosong Zhang +2Language Model Error DetectionConfidence Estimation in Language Models

  7. Look Before You Leap: Factual Decoding with Internal Attribution Signals

    Sep 14, 2026Hayeong Ryu, JungMin Yun, Byeonggeuk Lim +2Language Model DecodingLanguage Model Error Detection

  8. Domain-Specific Hallucination Detection in Large Language Models

    Sep 12, 2026Varun Teja Chundru, Debasmita BiswasLanguage Model Error DetectionLLM Hallucination Mitigation

  9. Legible Failures: Detecting and Repairing In-Context Binding Errors

    Sep 11, 2026Manas Venkata Sai Ravulapalli, Samrath Singh Chadha, Abhinav M. HariLanguage Model Error DetectionLLM Reliability

  10. Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs

    Aug 6, 2026Hamed Damirchi, Ignacio Meza De la Jara, Damith Ranasinghe +2Language Model Error DetectionLLM Interpretability

  11. UHP Detection: LVLMs have their Unique Hallucination Pattern in the Consistency Space

    Aug 4, 2026Amir Mohammad Ezzati, Kiyan Rezaee, Bardiya Kariminia +4Hallucination Detection in VLMsVision-Language Models

  12. D-Score: A Spectral Hidden-State Signal for Hallucination Detection in Large Language Models

    Jul 27, 2026Bianca Raimondi, Davide Evangelista, Maurizio Gabbrielli +1Language Model Error DetectionHallucination in Language Models

  13. Peeking Inside LLMs: Leveraging Internal Artifacts of LLMs for Enhancing Reliability in Legal Classification

    Jun 18, 2026Sudipta Santra, Debtanu Datta, Saptarshi GhoshLegal NLPLanguage Model Error Detection

  14. Measurement Under Selection: Decoy-Calibrated Failure Audits for Language Models

    Jun 8, 2026Vyzantinos Repantis, Ameya Gawde, Harshvardhan SinghLLM EvaluationLanguage Model Error Detection

  15. How Language Models Fail: Token-Level Signatures of Committed and Persistent Reasoning Failures

    Jun 4, 2026Tanvi Thoria, Kiana Jafari, Marc R. Schlichting +1Language Model Error DetectionLLM Reliability

  16. Ekka: Automated Diagnosis of Silent Errors in LLM Inference

    Jun 3, 2026Yile Gu, Zhen Zhang, Shaowei Zhu +4Software EngineeringLLM Serving

  17. TriLens: Per-Layer Logit-Lens Entropy for White-Box Hallucination Detection

    May 31, 2026Bohan Yang, Yijun Gong, Zhi Zhang +3Language Model Error DetectionLLM Reliability

  18. BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali

    May 29, 2026Shefayat E Shams Adib, Ahmed Alfey Sani, Ekramul Alam Esham +3Multilingual Language Model EvaluationLanguage Model Error Detection

  19. Learning the Error Patterns of Language Models

    May 27, 2026Jinwoo Kim, Taylor Berg-KirkPatrick, Loris D'AntoniConstrained DecodingLanguage Model Error Detection

  20. Entropy Distribution as a Fingerprint for Hallucinations in Generative Models

    May 27, 2026Mattia J. Villani, Pranav Deshpande, Akshay Seshadri +2Language Model Error DetectionLanguage Model Generation Evaluation

  21. Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

    May 26, 2026Prashanti Nilayam, Kiran Ramanna, Prashil TumbadeLLM Self-CorrectionLLM Evaluation

  22. CyberCorrect: A Cybernetic Framework for Closed-Loop Self-Correction in Large Language Models

    May 17, 2026Yuning Wu, Yingmin Liu, Yang ShuLLM Self-CorrectionLLM Evaluation

  23. A multilingual hallucination benchmark: MultiWikiQHalluA

    May 4, 2026Freja Thoresen, Dan Saattrup SmartMultilingual Language ModelsMultilingual Language Model Evaluation

  24. How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

    Apr 24, 2026Dharshan Kumaran, Viorica Patraucean, Simon Osindero +2LLM Self-CorrectionLanguage Model Error Detection

  25. HALT: Hallucination Assessment via Log-probs as Time series

    Feb 2, 2026Ahmad Shapiro, Karan Taneja, Ashok GoelLanguage Model Error DetectionHallucination Detection