LLM Abstention

LLM: Large Language Model

Momentum

11 papers in the last four weeks, up 120% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 61

All topics
CardsList
  1. TIPS: Topological Ill-Posedness Probing and Steering in Large Language Models

    Jun 22, 2026Guangyu Jiang, Sizhe Tang, Mahdi Imani +2Language Model SteeringQuestion Answering

  2. Coherence Under Commitment: Probing Generalization and Vacuous Memorization in LLM Logical Reasoning

    Jun 19, 2026Noor Islam S. Mohammad, Mahmudul HasanLLM EvaluationReasoning Consistency in Language Models

  3. DiagFlowBench: Evaluating How Language Models Handle Off-Procedure Inputs in Grounded Diagnostic Dialogue

    Jun 16, 2026Guillermo Gil de Avalle, Laura Maruster, Shaina Raza +1LLM EvaluationLLM Grounding

  4. PhantomBench: Benchmarking the Non-existential Threat of Language Models

    Jun 9, 2026Haeji Jung, Hila GonenLLM EvaluationHallucination in Language Models

  5. What Benchmarks Don't Measure: The Case for Evaluating Abstention Competence in Autonomous Agents

    Jun 1, 2026Victor Ojewale, Suresh VenkatasubramanianAgent EvaluationAI Agent Safety

  6. Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models

    May 30, 2026S M Tahmid Siddiqui, Akib Jawad Ononto, Anoop Singhal +1LLM Hallucination MitigationQuestion Answering

  7. OCC-RAG: Optimal Cognitive Core for Faithful Question Answering

    May 30, 2026Maksim Savkin, Mikhail Goncharov, Alexander Gambashidze +7Multi-Hop QALLM Grounding

  8. Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information

    May 27, 2026Renjie Gu, Jiaxu Li, Yihao Wang +8LLM AbstentionLarge Reasoning Models

  9. TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

    May 25, 2026Muyu Pan, Shu Zhao, Nan Zhang +4RL for Language ModelsHallucination Mitigation

  10. Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models

    May 25, 2026Ashwath Vaithinathan Aravindan, Mayank KejriwalConfidence Estimation in Language ModelsMultiple-Choice Question Answering

  11. Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

    May 24, 2026Chainarong AmornbunchornvejCognitive Architectures for AI AgentsComputational Cognitive Modeling

  12. Task Abstention for Large Language Models in Code Generation

    May 16, 2026Yanke Zhou, Yuhao Tan, Senrong Xu +4Selective PredictionLLM Hallucination Mitigation

  13. Quantifying and Mitigating Premature Closure in Frontier LLMs

    May 14, 2026Rebecca Handler, Suhana Bedi, Nigam ShahHealthcareLLM Prompting

  14. Trust or Abstain? A Self-Aware RAG Approach

    May 11, 2026Xi Zhu, Ziqi Wang, Kai Mei +5Retrieval-Augmented GenerationKnowledge Conflicts in Language Models

  15. CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine

    May 1, 2026Kevin H. Guo, Chao Yan, Avinash Baidya +5HealthcareLLM Reliability

  16. Geometry-Calibrated Conformal Abstention for Language Models

    Apr 30, 2026Rui Xu, Yi Chen, Sihong Xie +1Uncertainty QuantificationSelective Prediction

  17. Knowing When to Quit: A Principled Framework for Dynamic Abstention in LLM Reasoning

    Apr 20, 2026Hen Davidov, Nachshon Cohen, Oren Kalinsky +4RL for Language Model ReasoningLLM Abstention

  18. Learning from AVA: Early Lessons from a Curated and Trustworthy Generative AI for Policy and Development Research

    Apr 20, 2026Nimisha Karnatak, Mohamad Chatila, Daniel Alejandro Pinzón Hernández +3Retrieval-Augmented GenerationEvidence-Grounded Generation

  19. Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say "I Don't Know"

    Feb 4, 2026Dhruv Madhwal, Lyuxin David Zhang, Dan Roth +2LLM PromptingQuestion Decomposition

  20. Honesty over Accuracy: Trustworthy Language Models through Reinforced Hesitation

    Nov 14, 2025Mohamad Amin Mohamadi, Tianhao Wang, Zhiyuan LiRL for Language ModelsLLM Hallucination Mitigation

  21. Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety

    Oct 18, 2025Vamshi Krishna Bonagiri, Ponnurangam Kumaragurum, Khanh Nguyen +1LLM Agent EvaluationTool-Using Agents