LLM Self-Correction

LLM: Large Language Model

Latest papers 105

All topics
CardsList
  1. Agentic Molecular Recovery via Molecule-Aware Exploration

    Jun 4, 2026Suwan Yoon, Changhee LeeLLM Self-CorrectionMolecular Generation

  2. Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

    Jun 4, 2026Muhammad Talha Sharif, Abdul RehmanLLM Self-CorrectionMulti-Agent Reasoning

  3. Self-Reflective APIs: Structure Beats Verbosity for AI Agent Recovery

    Jun 3, 2026Arquimedes Canedo, Grama ChethanLLM Self-CorrectionTool-Use Evaluation

  4. Imbuing Large Language Models with Bidirectional Logic for Robust Chain Repair

    Jun 3, 2026Zehua Cheng, Wei Dai, Jiahao Sun +1LLM Self-CorrectionSymbolic Reasoning

  5. Learn from Your Mistakes: Tree-like Self-Play for Secure Code LLMs

    Jun 2, 2026Wenqi Chen, Ziyan Zhang, Bin Wang +3RL for Code GenerationLLM Self-Correction

  6. Med-HEAL: Analyzing and Mitigating Hallucinations in Medical LLMs with Hallucination-Aware In-Context Learning

    May 31, 2026Yiming Liao, Zeno Franco, Jose Eduardo Lizarraga Mazaba +1LLM Self-CorrectionLLM Hallucination Mitigation

  7. CRITIC-R1: Learning Structured Critics for Retrieval-Augmented Generation

    May 28, 2026Wenhan Xiao, Ziwei Zhang, Chuanyue Yu +4LLM Self-CorrectionRetrieval-Augmented Generation

  8. DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes

    May 27, 2026Caijun Xu, Changyi Xiao, Zhongyuan Peng +1LLM Self-CorrectionRL for Language Model Reasoning

  9. Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines

    May 26, 2026Prashanti Nilayam, Kiran Ramanna, Prashil TumbadeLLM Self-CorrectionLLM Evaluation

  10. ProCrit: Self-Elicited Multi-Perspective Reasoning with Critic-Guided Revision for Multimodal Sarcasm Detection

    May 20, 2026Yingjia Xu, Jiulong Wu, Bowen Zhang +3LLM Self-CorrectionMultimodal CoT Reasoning

  11. CyberCorrect: A Cybernetic Framework for Closed-Loop Self-Correction in Large Language Models

    May 17, 2026Yuning Wu, Yingmin Liu, Yang ShuLLM Self-CorrectionLLM Evaluation

  12. Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards

    May 14, 2026Mengjie Ren, Jie Lou, Boxi Cao +6RL for Code GenerationLLM Self-Correction

  13. Your Simulation Runs but Solves the Wrong Physics: PDE-Grounded Intent Verification for LLM-Generated Multiphysics Simulation Code

    May 10, 2026Zhenghan Song, Yulong Liu, Cheng Wan +4LLM Self-CorrectionScientific Code Generation

  14. Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories

    May 9, 2026Dongcheng Zhang, Yi Zhang, Yuxin Chen +3LLM Self-CorrectionRL for Language Model Reasoning

  15. Knowing but Not Correcting: Routine Task Requests Suppress Factual Correction in LLMs

    May 7, 2026Zixuan Chen, Hao Lin, Zizhe Chen +6LLM Self-CorrectionLLM Reliability

  16. The Counterexample Game: Iterated Conceptual Analysis and Repair in Language Models

    May 5, 2026Daniel Drucker, Kyle MahowaldLLM Self-CorrectionLLM Evaluation

  17. Revisiting the Travel Planning Capabilities of Large Language Models

    May 5, 2026Bo-Wen Zhang, Jin Ye, Peng-Yu Hua +4LLM Self-CorrectionLLM Evaluation

  18. GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair

    May 1, 2026Yinhao Xiao, Rongbo Xiao, Yihan ZhangLLM Self-CorrectionAutomated Program Repair

  19. LLMs as ASP Programmers: Self-Correction Enables Task-Agnostic Nonmonotonic Reasoning

    Apr 30, 2026Adam Ishay, Joohyung LeeLLM Self-CorrectionLogical Reasoning

  20. OmniVTG: A Large-Scale Dataset and Training Paradigm for Open-World Video Temporal Grounding

    Apr 28, 2026Minghang Zheng, Zihao Yin, Yi Yang +2Temporal Video GroundingLLM Self-Correction

  21. Faithful Autoformalization via Roundtrip Verification and Repair

    Apr 27, 2026Daneshvar Amrollahi, Jerry Lopez, Clark BarrettLLM Self-CorrectionLegal NLP

  22. When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL2DL

    Apr 25, 2026Yijiashun Qi, Xiang Xu, Yuxuan LiLLM Self-CorrectionLogical Reasoning

  23. How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals

    Apr 24, 2026Dharshan Kumaran, Viorica Patraucean, Simon Osindero +2LLM Self-CorrectionLanguage Model Error Detection

  24. Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

    Apr 21, 2026Clara Lachenmaier, Hannah Bultmann, Sina ZarrießLLM Self-CorrectionHuman-AI Interaction

  25. Decompose, Structure, and Repair: A Neuro-Symbolic Framework for Autoformalization via Operator Trees

    Apr 21, 2026Xiaoyang Liu, Zineng Dong, Yifan Bai +3LLM Self-CorrectionNeuro-Symbolic Reasoning