cs.AISep 28, 2026

SkillFocus: Evolving Agent Skills via Capability Decomposition

Authors: Ning Wang, Zhiren Gong, Bingdong Li, Peng Yang, Aimin Zhou

Organizations: East China Normal University · Nanyang Technological University · Southern University of Science and Technology · Shanghai Innovation Institute

Abstract

Agent skill evolution seeks to improve reusable procedural guidance for large language model (LLM) agents through iterative revision. Existing methods base each revision mainly on execution trajectories or feedback, leaving recurring behavioral requirements across tasks implicit and tying revision to the behavior of the current skill. We introduce SkillFocus, which decomposes recurring task requirements into a capability space that remains fixed as the skill evolves, separating what tasks require from how the current skill behaves. SkillFocus maps current task outcomes to this space to identify the capability that leaves the most tasks unresolved, then uses that capability to determine what to revise and which evidence to use. Across four benchmarks spanning heterogeneous tasks, SkillFocus achieves the best held-out accuracy on all four, outperforming the strongest competing result by 5.7 points on average while using 24% fewer evolution tokens on average than the closest iterative baseline. Controlled studies further show that capabilities derived from recurring task requirements outperform task-semantic and execution-derived alternatives, while randomizing task--capability assignments reduces final accuracy by up to 20.2 points. Matching evidence to the selected capability increases candidate gain by 4.4 points under prioritized revision.

Figures & tables

Appendix figures & tables17 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. SkillAxe: Sharpening LLM-Authored Agent Skills Through Evaluation-Guided Self-Refinement

    Jun 9, 2026Srishti Gautam, Arjun Radhakrishna, Sumit GulwaniSkillsInstruction

  2. SkillGrad: Optimizing Agent Skills Like Gradient Descent

    May 26, 2026Hanyu Wang, Yifan Lan, Bochuan Cao +2SkillsLarge Language Model Agents