Language Model Self-Assessment

Latest papers 56

All topics
CardsList
  1. Identifying Introspection From the Inside

    Oct 5, 2026David I. Atkinson, Dillon Plunkett, David BauLLM InterpretabilityLanguage Model Introspection

  2. Also Small Models Can Reasonably Self-Evaluate Their Confidence

    Sep 30, 2026Idil Kapikiran, Thomas Decker, Thomas RunklerConfidence Estimation in Language ModelsSmall Language Models

  3. Audio LLMs Know When They Can't Hear You

    Sep 24, 2026Amirhosein Javadi, Richa Dixit, Mehrdad Farajtabar +3Audio-Language Model EvaluationLanguage Model Self-Assessment

  4. When Should LLMs Abstain? Chain-of-Self-Questioning for Selective Risk Control

    Sep 15, 2026Ali ŞenolSelective PredictionLLM Prompting

  5. The Assistant's Ideal Self

    Aug 31, 2026Mert YazanPersonality Modeling in Language ModelsLanguage Model Self-Assessment

  6. Evaluating and Improving LLM Self-Modeling

    Aug 31, 2026Siqi Zeng, Andre N. Assis, Rowan WangLLM EvaluationLanguage Model Introspection

  7. Reinforcing Step-level Reasoning for Effective Self-Correction in LLMs

    Aug 12, 2026Vu Duc Anh, Nhat M. Hoang, Do Xuan Long +3LLM Self-CorrectionDirect Preference Optimization

  8. The Personalization Mirage: How LLMs Fabricate User Profiles, and Why Self-Monitoring Misleads

    Aug 5, 2026Yushi Sun, Yanjie Zhang, Rui ShengLLM EvaluationLLM Personalization

  9. How Much Does a Reasoning Summary Reveal? An Observability Ladder for Large Language Models

    Aug 3, 2026Andres Algaba, Francesca Carlon, Lynn Delcon +3LLM EvaluationLLM Interpretability

  10. Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

    Jul 31, 2026Yefan Tao, Gerald Friedland, Madhusudhanan Chandrasekaran +1LLM Self-CorrectionLLM Self-Refinement

  11. SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute

    Jul 30, 2026Hongyu Chen, Liang Lin, Guangrun WangRL for Language Model ReasoningLanguage Model Self-Assessment

  12. One Human, NN Agents: Audit-Budget Allocation for LLM Agent Fleets under Miscalibrated, Correlated Confidence

    Jul 30, 2026Cesare Zavattari, Alessandro Tommasi, Giuseppe PrencipeScalable OversightAI Agent Reliability

  13. Reality Monitoring in Large Language Models: Self-Knowledge That Transforms with Conversation Memory

    Jul 27, 2026Saurabh Ranjan, Konstantina Sokratous, Brian OdegaardHallucination in Language ModelsLanguage Model Self-Assessment

  14. The Two-Process Theory of Machine Self-Report

    Jul 22, 2026Hubert Plisiecki, Filip Chmielewski, Kacper Dudzic +3Personality Modeling in Language ModelsLanguage Model Self-Assessment

  15. FIFA World Cup 2026 as a Contamination-Free Benchmark for LLM Forecasting Agents: Four Models, a Bookmaker, and 104 Matches

    Jul 20, 2026Jiacheng Ding, Cong Guo, Jason XuLLM Agent EvaluationForecasting Benchmarks

  16. Diagnosing Correctness Probes under Self-Judgement Confounding

    Jul 18, 2026Yi-Long LuLLM InterpretabilityLanguage Model Self-Assessment

  17. Metacognition in LLMs: Foundations, Progress, and Opportunities

    Jul 13, 2026Gabrielle Kaili-May Liu, Areeb Gani, Jacqueline Lu +3Language Model Self-AssessmentMetacognition in Language Models

  18. Revealing Hidden Model Behaviors with Task-Specific Self-Reports

    Jul 3, 2026Taras Kutsyk, Bartosz ZielińskiLLM AuditingLanguage Model Introspection

  19. Latent Confidence Alignment for LLM Self-Assessment

    Jun 20, 2026Ting-Yu Chen, Tingting Yu, Pei-Cing Huang +3LLM EvaluationConfidence Estimation in Language Models

  20. Self-Evaluation Is Already There: Eliciting Latent Judge Calibration in Base LLMs with Minimal Data

    Jun 3, 2026XiuYu Zhang, Yi Shan, Junfeng Fang +1LLM EvaluationLLM-as-a-Judge

  21. CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

    Jun 2, 2026Alexander Apartsin, Yehudit ApersteinModel SelectionLLM Evaluation

  22. Capability Self-Assessment: Teaching LLMs to Know Their Limits

    May 29, 2026Haoyan Yang, Reza Shirkavand, Yukai Jin +3LLM EvaluationLanguage Model Self-Assessment

  23. What Am I Missing? Question-Answering as Hidden State Probing

    May 29, 2026Chu Fei Luo, Samuel Dahan, Xiaodan ZhuLLM Self-RefinementLanguage Model Self-Assessment

  24. Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs

    May 19, 2026Carolina Camassa, Derek ShillerLLM EvaluationLanguage Model Self-Assessment

  25. Beyond Confidence: Rethinking Self-Assessments for Performance Prediction in LLMs

    May 8, 2026Sree Bhattacharyya, Samarth Khanna, Leona Chen +3LLM EvaluationConfidence Estimation in Language Models

  26. EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics

    May 5, 2026Shuyue Stella Li, Rui Xin, Teng Xiao +8RL for Language ModelsLLM Evaluation

  27. Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty

    Apr 19, 2026Jingyi Ren, Ante Wang, Yunghwei Lai +5LLM EvaluationSelective Prediction

  28. Introspection Adapters: Training LLMs to Report Their Learned Behaviors

    Apr 18, 2026Keshav Shenoy, Li Yang, Abhay Sheshadri +4LLM AuditingLanguage Model Introspection

  29. Are LLM Evaluators Really Narcissists? Sanity Checking Self-Preference Evaluations

    Jan 30, 2026Dani Roytburg, Matthew Bozoukov, Matthew Nguyen +3LLM-as-a-JudgeLanguage Model Generation Evaluation

  30. No Reliable Evidence of Self-Reported Sentience in Small Large Language Models

    Jan 20, 2026Caspar Kaiser, Sean EnderbyLLM EvaluationLanguage Model Self-Assessment

  31. Calibration Is Not Enough: Evaluating Confidence Estimation Under Language Variations

    Jan 12, 2026Yuxi Xia, Dennis Ulmer, Terra Blevins +3LLM EvaluationConfidence Estimation in Language Models