LLM Self-Correction

LLM: Large Language Model

Latest papers 105

All topics
CardsList
  1. Can Language Models Learn to Reject Their Own Bad Reasoning Steps?

    Oct 5, 2026Siheng Xiong, Xiaoze Liu, Yiqiao Jin +2LLM Self-CorrectionLanguage Model Decoding

  2. Self-Spec Verifiable Code Generation

    Sep 30, 2026Jiaru Qian, Yihong Dong, Yongmin Li +3LLM Self-CorrectionCode Generation

  3. After the Fix: Transfer of Corrected Agent Experience

    Sep 28, 2026Yanfei Zhang, Xu LinLLM Self-CorrectionLLM Agent Memory

  4. Exact Feedback Is Not Control: Evaluating Text-based Closed-Loop Revision in LLMs

    Sep 23, 2026Haitong Jiang, Chunlin Liu, Yile Wang +1LLM Self-CorrectionControllable Text Generation

  5. Sage: Formalization with Semantic Correction

    Sep 16, 2026Thomas Hirtz, Farzad Jafarrahmani, Abdelmouksit Sagueni +3LLM Self-CorrectionLean Theorem Proving

  6. RetroThinker: Enabling Retrospective Thinking in Speech LLMs

    Sep 12, 2026Yi-Jen Shih, Puyuan Peng, Abdelrahman Mohamed +1LLM Self-CorrectionEfficient Language Model Reasoning

  7. If It's Not Buggy, Don't Fix It: On the Dynamics of Iterative Bug-fixing with LLMs

    Sep 9, 2026Xietao Wang-Lin, Anton Isopoussu, Louis MahonLLM Self-CorrectionAutomated Program Repair

  8. Towards Expert Financial QA via Self-Improving RAG

    Aug 27, 2026Junjie Xiong, Shawheen Ghezavat, Aum HirparaLLM Self-CorrectionFinancial QA

  9. Reinforcing Step-level Reasoning for Effective Self-Correction in LLMs

    Aug 12, 2026Vu Duc Anh, Nhat M. Hoang, Do Xuan Long +3LLM Self-CorrectionDirect Preference Optimization

  10. Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning

    Aug 6, 2026Ahsan Bilal, Muhammad Ahmed Mohsin, Muhammad Umer +4Numerical Reasoning in Language ModelsLLM Self-Correction

  11. AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction

    Jul 31, 2026Rui Zou, Yutao Zhu, Mengqi Wei +1LLM Self-CorrectionLLM Answer Verification

  12. Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

    Jul 31, 2026Yefan Tao, Gerald Friedland, Madhusudhanan Chandrasekaran +1LLM Self-CorrectionLLM Self-Refinement

  13. GGC: Selective Query Correction for Reliable Text-to-SPARQL Generation

    Jul 30, 2026Ziyi Yang, Thanh-Son Nguyen, Tuan Anh Nguyen +1LLM Self-CorrectionLLM Reliability

  14. LA-RL: Label-Aware Self-Reflection for Reinforcement Learning in Information Extraction

    Jul 26, 2026Xiao You, Tianwei Yan, Zixu Shan +2LLM Self-CorrectionRL for Language Models

  15. PhoenixRepair: Rethinking Repair Strategy Exploration in Software Agents

    Jul 21, 2026Tianyue Jiang, Yanlin Wang, Xin He +7LLM Self-CorrectionSoftware Engineering Agents

  16. Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents

    Jul 20, 2026Yitao Wu, Si Shen, Rui Yang +2AI Agent ReliabilityLLM Self-Correction

  17. Grounded verification of chemical and materials reasoning: detection is the bottleneck

    Jul 19, 2026Can Polat, Mustafa Kurban, Erchin Serpedin +1LLM Self-CorrectionLLM Grounding

  18. Reward-Driven LLM Agent Workflows: Synthesizing POMDP Routing and Self-Correction for Autonomous Decision-Making

    Jul 19, 2026Amez Amanj Ali, Kuo-Kun TsengReward ModelingLLM Self-Correction

  19. Deep Interaction: An Efficient Human-AI Interaction Method for Large Reasoning Models

    Jul 15, 2026Hefeng Zhou, Jinxuan Zhang, Jiong Lou +4LLM Self-CorrectionLLM Prompting

  20. Experience Memory Graph: One-Shot Error Correction for Agents

    Jul 15, 2026Wenjun Wang, Yuchen Fang, Fengrui Liu +2Agent MemoryLLM Self-Correction