cs.AIOct 8, 2026

Overcoming Prior Barriers: Supervised Fine-Tuning under Long-Tail Distribution

Authors: Haohui Wang, Jiahao Xu, Wangzhi Zhan, Tong Zeng, Dongqi Fu, Hong Li, Swastik Roy, Naren Ramakrishnan, +4 more

Organizations: Virginia Tech · Amazon · Meta · MBZUAI · Dartmouth College

Abstract

Supervised fine-tuning (SFT) adapts pretrained large language models (LLMs) to downstream tasks, but the required concepts can receive substantially different levels of pretrained support. Frequent concepts are more likely to be well learned, whereas rare concepts may remain weakly represented. We introduce a novel notion named prior barrier to quantify how strongly the pretrained model supports competing concepts over the target concept. We observe that prior barriers follow a long-tail distribution, placing head and tail concepts at different starting points for SFT: head concepts face lower prior barriers, whereas tail concepts require additional instructions to overcome their higher prior barriers. Our theoretical analysis further derives a predictive risk bound for SFT under long-tail prior barriers, explicitly characterizing how the prior barrier and accumulated SFT evidence jointly determine predictive performance. Motivated by this prior barrier-dependent demand, we propose PASS, an adaptive SFT instruction selection method that constructs reference-derived concepts and estimates the distinguishing evidence provided by each instruction, and adaptively allocates the selection budget toward concepts that remain insufficiently covered under the current selection. In this way, PASS jointly considers which instructions can provide useful evidence and where additional supervision is needed under a limited budget. Experiments show that our method consistently outperforms seven state-of-the-art instruction selection methods on four backbone-budget settings. An ablation study further shows that PASS's adaptive allocation consistently improves over uniform allocation.

Figures & tables

Appendix figures & tables2 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. PriFT: Prior-Support Guided Supervised Fine-Tuning

    Jun 8, 2026Ke Wang, Shuangqi Li, Mathieu Salzmann +1Supervised Fine-TuningFine-Tuning

  2. LP-SFT: Local-Preserving Supervised Fine-Tuning via Multimodal Entropy Structure

    Jul 6, 2026Yueyang Wang, Baolong Bi, Shuo Lu +1Supervised Fine-TuningFine-Tuning

  3. A Unifying Lens on Supervised Fine-Tuning Through Target Distribution Design

    Jun 9, 2026Tong Xie, Yuanhao Ban, Yunqi Hong +3Supervised Fine-TuningFine-Tuning