cs.CLApr 22, 2026

RADS: Reinforcement Learning-Based Sample Selection Improves Transfer Learning in Low-resource and Imbalanced Clinical Settings

Authors: Wei HanDavid MartinezAnna KhaninaLawrence CavedonKarin Verspoor

Organizations: School of Computing Technologies, RMIT University · Department of Infectious Disease, Peter MacCallum Cancer Centre · 3National Centre for Infections in Cancer, Melbourne · 5Sir Peter MacCallum Department of Oncology, The University of Melbourne · School of Computing and Information Systems, The University of Melbourne

Abstract

A common strategy in transfer learning is few shot fine-tuning, but its success is highly dependent on the quality of samples selected as training examples. Active learning methods such as uncertainty sampling and diversity sampling can select useful samples. However, under extremely low-resource and class-imbalanced conditions, they often favor outliers rather than truly informative samples, resulting in degraded performance. In this paper, we introduce RADS (Reinforcement Adaptive Domain Sampling), a robust sample selection strategy using reinforcement learning (RL) to identify the most informative samples. Experimental evaluations on several real world clinical datasets show our sample selection strategy enhances model transferability while maintaining robust performance under extreme class imbalance compared to traditional methods.

Explore similar work

CardsList