LLM Agent Self-Improvement

LLM: Large Language Model

Momentum

44 papers in the last four weeks, up 267% on the four weeks before. 0.4% of all new papers.

Jul 13Week of Sep 28

Latest papers 189

All topics
CardsList
  1. Fantastic Scientific Agents and How to Build Them: AgentBuild for Rietveld Refinement

    Jun 11, 2026Woong Shin, Craig A. Bridges, Marshall T. McDonnell +1AI Agents for Scientific DiscoveryLLM Agent Self-Improvement

  2. EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents

    Jun 9, 2026Weixian Xu, Shilong Liu, Mengdi WangContinual Learning for LLM AgentsLLM Agent Self-Improvement

  3. Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

    Jun 9, 2026Xucong Wang, Ziyu Ma, Shidong Yang +4LLM Agent Self-ImprovementLLM Agent Training

  4. OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

    Jun 8, 2026Mingxian Lin, Shengju Qian, Yuqi Liu +9VLM EvaluationGame-Playing Agents

  5. Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

    Jun 8, 2026Rishabh Sabharwal, Hongru Wang, Amos Storkey +1Deep Research AgentsAI Agent Evaluation

  6. From 0-to-1 to 1-to-N: Reproducible Engineering Evidence for MetaAI Recursive Self-Design

    Jun 8, 2026Dun Li, Jiatao Li, Hongzhi LiLLM Agent Self-ImprovementSelf-Improving Agents

  7. Self-Harness: Harnesses That Improve Themselves

    Jun 8, 2026Hangfan Zhang, Shao Zhang, Kangcong Li +5Agent Harness OptimizationLLM Agent Self-Improvement

  8. Bayesian-Agent: Posterior-Guided Skill Evolution for LLM Agent Harnesses

    Jun 6, 2026Xiaojun Wu, Cehao Yang, Honghao Liu +7Agent Harness OptimizationLLM Agent Workflow Optimization

  9. SkillComposer: Learning to Evolve Agent Skills for Specification and Generalization

    Jun 4, 2026Qi Zhang, Zhaopeng Feng, Xiaonan Shi +8LLM Agent Self-ImprovementLLM Agent Skill Learning

  10. Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference

    Jun 4, 2026Wenbo Pan, Shujie Liu, Chin-Yew Lin +5Software Engineering AgentsAgent Harness Optimization

  11. SePO: Self-Evolving Prompt Agent for System Prompt Optimization

    Jun 3, 2026Wangcheng Tao, Han Wu, Weng-Fai WongLLM Agent Self-ImprovementAutomatic Prompt Optimization

  12. FederatedSkill: Federated Learning for Agentic Skill Evolution

    Jun 2, 2026Jingbo Yang, Guanyu Yao, Yang Zhang +3LLM Agent Self-ImprovementLLM Agent Skill Learning

  13. Better with Experience: Self-Evolving LLM Agents for Evidence-Grounded Health Community Notes

    Jun 1, 2026Zihang Fu, Fanxiao Li, Jianyang Gu +5HealthcareLLM Agent Self-Improvement

  14. HarnessForge: Joint Harness and Policy Evolution for Adaptive Agent Systems

    Jun 1, 2026Mingju Chen, Can Lv, Guibin Zhang +2Agent Harness OptimizationLLM Agent Self-Improvement

  15. Adaptive Auto-Harness: Sustained Self-Improvement for Agentic System Deployment on Open-Ended Task Streams

    Jun 1, 2026Zewen Liu, Zhan Shi, Yisi Sang +7Continual Learning for LLM AgentsAgent Harness Optimization

  16. SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision

    May 31, 2026Yuxuan Liu, Zhaochen Su, Lingyun Xie +11LLM Agent Self-ImprovementLLM Agent Skill Learning

  17. AutoSci: A Memory-Centric Agentic System for the Full Scientific Research Lifecycle

    May 29, 2026Weitong Qian, Beicheng Xu, Zhongao Xie +16Multi-Agent LLM SystemsLLM Agent Memory

  18. MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

    May 29, 2026Bowen Ren, Heyan Huang, Yinghao Li +1Continual Learning for LLM AgentsLLM Agent Self-Improvement

  19. Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

    May 28, 2026Minhua Lin, Juncheng Wu, Zijun Wang +14LLM Agent EvaluationAgent Harness Optimization

  20. GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents

    May 28, 2026Johannes Moll, Jean-Philippe Corbeil, Jiazhen Pan +4LLM Agent Self-ImprovementLLM Agent Skill Learning

  21. SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents

    May 28, 2026Wentao Hu, Zhendong Chu, Yiming Zhang +6LLM Agent Self-ImprovementLLM Agent Skill Learning

  22. SkillGrad: Optimizing Agent Skills Like Gradient Descent

    May 26, 2026Hanyu Wang, Yifan Lan, Bochuan Cao +2Large Language Model-Guided OptimizationLLM Agent Self-Improvement

  23. SIA: Self Improving AI with Harness & Weight Updates

    May 26, 2026Prannay Hebbar, Yogendra Manawat, Samuel Verboomen +4Agent Harness OptimizationLLM Agent Self-Improvement

  24. CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly

    May 25, 2026Yihe Fan, Changyi Li, Lichen Xu +4LLM Agent Self-ImprovementAutomated Penetration Testing

  25. SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository Setup?

    May 25, 2026Zihang Zhou, Ziqian Ren, Yukai Wu +7Software Engineering AgentsLLM Agent Self-Improvement