cs.CLOct 1, 2026

Scalable, Transferable Meta-network for Data Selection Requires a Different Loss (and Why the Obvious Choice is Problematic)

Authors: Zilin Du, Bowen Yang, Boyang Albert Li

Organizations: College of Computing and Data Science Nanyang Technological University Singapore

Abstract

Data selection is critical for training large language models on massive and heterogeneous corpora. Meta-learning for Training-data Selection offers a principled alternative to heuristic scoring by learning data weights from a target validation objective, but existing methods face a trade-off between fine-grained valuation and transferability to unseen data. A natural solution is to replace per-sample weights with a selection network. However, we find that directly incorporating such a network into existing MTS objectives leads to unstable optimization and poor generalization, caused by weight suppression and persistent reliance on easy-to-learn features. To address these issues, we propose Transferable Example Scoring and Selection (TESS), a scalable data-selection framework built on a Pointwise Value Matching objective (PVM). Experiments on LLM safety and targeted instruction tuning demonstrate strong transfer across datasets, from subsets to full corpora, and from smaller to larger models.

Figures & tables

Appendix figures & tables1 asset

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training

    May 8, 2026Pingbang Hu, Xueshen Liu, Z. Morley Mao +1Training DataRegularization

  2. The Long-Term Effects of Data Selection in LLM Fine-Tuning

    May 28, 2026Yuxin Yang, Aoxiong Zeng, Xiangquan YangLarge Language Model Fine-TuningModel Fine-Tuning