LLM Reliability

LLM: Large Language Model

Momentum

78 papers in the last four weeks, up 105% on the four weeks before. 0.8% of all new papers.

Jul 13Week of Sep 28

Latest papers 507

All topics
CardsList
  1. Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models

    Jun 1, 2026Simone Caldarella, Davide Talon, Rahaf Aljundi +2LLM EvaluationOverthinking in Language Models

  2. Does Compression Preserve Uncertainty? A Unified Benchmark for Quantized and Sparse LLMs via Conformal Prediction

    Jun 1, 2026Yujia Tong, Yuxi Wang, Yunyang Wan +3LLM QuantizationLLM Evaluation

  3. TriLens: Per-Layer Logit-Lens Entropy for White-Box Hallucination Detection

    May 31, 2026Bohan Yang, Yijun Gong, Zhi Zhang +3Language Model Error DetectionLLM Reliability

  4. Accuracy, Stability, and Repeated-Run Reliability of Large Language Models on Deterministic Programming Tasks

    May 30, 2026Yongxi Zhou, Lai Yun Choi, Jiaxi Wen +1LLM EvaluationLanguage Model Generation Evaluation

  5. Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models

    May 30, 2026S M Tahmid Siddiqui, Akib Jawad Ononto, Anoop Singhal +1LLM Hallucination MitigationQuestion Answering

  6. Sequential statistical inference for Large Language Models: Representation, validity, and monitoring

    May 30, 2026Yao XieLanguage Model CalibrationLLM Reliability

  7. Capability Self-Assessment: Teaching LLMs to Know Their Limits

    May 29, 2026Haoyan Yang, Reza Shirkavand, Yukai Jin +3LLM EvaluationLanguage Model Self-Assessment

  8. BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali

    May 29, 2026Shefayat E Shams Adib, Ahmed Alfey Sani, Ekramul Alam Esham +3Multilingual Language Model EvaluationLanguage Model Error Detection

  9. LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories

    May 29, 2026Krishnapriya Vishnubhotla, Sowmya Vajjala, Akriti Vij +1Multilingual Language Model EvaluationLLM-as-a-Judge

  10. LLM-FACETS: A Privacy-Preserving Framework for Evaluating LLM Transparency and Accountability

    May 29, 2026Tom Lucas, Alessio Buscemi, Alfredo Capozucca +2LLM EvaluationAI Accountability

  11. Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits

    May 29, 2026Soorya Ram Shimgekar, Agam Goyal, Amruta Parulekar +6Prompt SensitivityLLM Interpretability

  12. The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability

    May 28, 2026Mikhail L. Arbuzov, Lee Mosbacker, Sisong Bei +3LLM Reliability

  13. SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

    May 28, 2026Sy-Tuyen Ho, Minghui Liu, Huy Nghiem +1AI for ScienceLLM Evaluation

  14. Harnessing non-adversarial robustness in large language models

    May 28, 2026Qinghua Zhou, Ellina Aleshina, Andrey Lovyagin +6LLM Fine-TuningNeural Network Robustness

  15. Entropy Distribution as a Fingerprint for Hallucinations in Generative Models

    May 27, 2026Mattia J. Villani, Pranav Deshpande, Akshay Seshadri +2Language Model Error DetectionLanguage Model Generation Evaluation

  16. The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages

    May 27, 2026Eric Onyame, Runtao Zhou, Kowshik Thopalli +2CoT FaithfulnessMultilingual Language Model Evaluation

  17. Chain-based Adaptive Reconfiguration Over Lattices for Hallucination Reduction

    May 26, 2026Joan Vendrell Gallart, Solmaz Kia, Russell Bent +1LLM Hallucination MitigationHallucination in Language Models

  18. Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination

    May 26, 2026Yedidia Agnimo, Anna Korba, Annabelle Blangero +2LLM EvaluationHallucination in Language Models

  19. Innovation: An Almost Characterization of Hallucination

    May 26, 2026Nishant P. Das, Piyush SrivastavaHallucination in Language ModelsLanguage Model Calibration

  20. SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation Reliability

    May 26, 2026Ramakrishna Vamsi Setti, Jagadeesh Rachapudi, Sachin Chaudhary +2LLM ReliabilityLanguage Model Robustness

  21. The Attribution Blind Spot: Detecting When Language Models Rely on Memory Rather Than Retrieved Context

    May 26, 2026Zhe Yu, Wenpeng Xing, Yunzhao Wei +4Memorization in Language ModelsLLM Interpretability

  22. LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

    May 25, 2026Lingyao Li, Junjie Xiong, Changjia Zhu +5LLM EvaluationLLM-as-a-Judge