cs.MAJul 20, 2025

Knowing When to Critique: Task-Adaptive Metacognitive Regulation for Reliable LLM Reasoning

Authors: Xinmeng Hou, Ziting Chang, Zhouquan Lu, Bohao Qu, Liang Wan, Wei Feng, Hai Hu, Qing Guo

Organizations: Nanyang Technological University · Unicorn Verse · Shanghai Jiao Tong University · Tianjin University · City University of Hong Kong · Nankai University

Abstract

Large language models (LLMs) reason fluently but do not regulate their reasoning: they apply uniform scrutiny to every input, which leaves them vulnerable to adversarial and counterfactual prompts, while indiscriminate critique over-corrects answers that were already sound. We propose MetaCrit, a multi-agent framework grounded in Nelson and Narens' metacognitive regulation theory that calibrates how much critique each task receives. MetaCrit separates regulation into four agents: an object-level generator, a monitoring agent that assesses response validity, a control agent that critiques logical soundness, and a meta-level synthesizer that reconciles their signals into a regulated response. Adaptivity here is input-conditioned intervention strength within a fixed pipeline: all four agents run on every input and what varies is the direction and magnitude of the correction they produce, not which stages execute. Across reasoning, safety, and bias benchmarks, MetaCrit improves truthfulness and logical soundness and reaches zero toxicity on BOLD and HONEST without a reasoning trade-off, whereas the same critique applied indiscriminately degrades performance. The cost is four calls per query, about one sixth of the cost of a dedicated reasoning model of similar accuracy. A writing study shows that MetaCrit is preferred for critical-thinking support, and its agents transfer to existing frameworks without architectural change. Code is available at https://github.com/Paparare/EduThink4AI.

Figures & tables

Appendix figures & tables45 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Cognitive Demand Steering for Adaptive Meta-Reasoning in Large Language Models

    Aug 2, 2026John Scoville, Shengzhuang Chen, Yejin Bang +2LLM Reasoning StrategiesCognitive Science

  2. Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

    May 22, 2026Sirui Chen, Lei Xu, Yuying Zhao +6LLM Reasoning StrategiesMetacognition

  3. Decomposing and Steering Functional Metacognition in Large Language Models

    May 9, 2026Yanshi Li, Xueru Bai, Shuman Liu +2MetacognitionLLM Reasoning Strategies