cs.AISep 27, 2026

COEVO: Co-Evolving Context and Parameters for Recursive Self-Improvement

Authors: Siwei Chen, Xinping Bao, Xinyu Cai, Yuan Cao, Wan Jiang, Shaohong Chen

Organizations: Peking University · Emotional Machine

Abstract

Recursive self-improvement (RSI) seeks to move large language models beyond static training pipelines toward systems that can participate in improving their own future behavior. Existing approaches largely follow two directions: updating model parameters through online learning, or improving the external context through search, reflection, and prompt optimization. Although both mechanisms can support continued improvement, they are typically studied independently. This separation overlooks an important interaction: the context shapes the experience from which a model learns, while an evolving model may interpret and utilize the same context differently over time. We therefore formulate RSI as a problem of parameter--context co-evolution, where model parameters and the learning context adapt within a shared feedback loop. We introduce COEVO, a framework that updates model parameters from on-policy experience while adapting contextual guidance according to the state of the evolving policy. Policy entropy and prompt-conditioned attention are used as complementary signals to guide this adaptation. Experiments show that COEVO consistently improves task performance over fixed-context reinforcement learning and produces policies that are more robust to changes in system prompts. More broadly, our results suggest that external context should be viewed not merely as a fixed interface to a large language model, but as an adaptive component of recursive self-improvement.

Figures & tables

Appendix figures & tables8 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Learning What to Remember and What to Internalize in LLM Self-Evolution via Adaptive Memory-Parameter Coordination

    Aug 2, 2026Tianyun Ji, Zhenya Huang, Jiayu Liu +3Self-Evolving AgentsSelf-Evolution

  2. Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

    May 27, 2026Bowen Wei, Nan Wang, Yuqing Zhou +2LLM Reasoning StrategiesSelf-Evolution

  3. Self-Improving Large Language Models via Progressive Experience Evolution

    Aug 3, 2026Shijie Ren, Xiting Wang, Meng Li +8Unsupervised On-Policy Self-DistillationOn-Policy Self-Evolution