cs.AIOct 1, 2026

PG-SFT: Balancing Capability Acquisition and Retention in Offline Agent Fine-Tuning

Authors: Ronghua Li, Zi Liang, Zhishan Li, Shinan Liu

Organizations: The University of Hong Kong · The Hong Kong Polytechnic University · Independent Researcher

Abstract

Supervised fine-tuning (SFT) on offline agent trajectories is the standard approach for training specialized tool-using agents, but forcing models to imitate reasoning and actions token by token may harm other capabilities (e.g., general reasoning, tool calling, code generation) of the base model. In this work, we focus on studying \emph{how to better balance the trade-off between acquiring new capabilities and preserving existing ones during agent trace SFT}. By comparing several baselines in our setup, standard SFT improves the target benchmark while lowering several non-target benchmark scores; meanwhile, simply constraining distributional drift using KL penalty or limiting the update magnitude did not avoid this regression trend. Motivated by recent token-wise adaptive learning objectives, this work proposes \textbf{Privilege-Guided SFT (PG-SFT)} to leverage turn-level information gain of agent trajectories as an indicator to adjust supervision strength. PG-SFT yields a more favorable observed trade-off on the evaluated benchmarks, substantially reducing distributional drift and broad capability degradation at the cost of slight degradation in target-task performance. Our findings suggest that balancing the acquisition--retention trade-off depends not only on whether the model is anchored to its base behavior, but also on where and how strongly supervision should depart from that behavior.}

Figures & tables

Appendix figures & tables2 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. TRACE: Capability-Targeted Agentic Training

    Apr 7, 2026Hangoo Kang, Tarun Suresh, Jon Saad-Falcon +1Agentic LearningSelf-Improving Agents

  2. SFT or RL for Tool-Calling Agents? A Controlled Study Across Data, Method, and Scale

    Sep 15, 2026Md Tahmid Rahman Laskar, Xue-Yong Fu, Shashi Bhushan TNLanguage-Model AgentsFlow-Grpo

  3. PriFT: Prior-Support Guided Supervised Fine-Tuning

    Jun 8, 2026Ke Wang, Shuangqi Li, Mathieu Salzmann +1Model Fine-Tuning