cs.CVOct 28, 2024

Large Pretraining Datasets Don't Guarantee Robustness after Fine-Tuning in Image Classification

Authors: Jaedong Hwang, Brian Cheung, Zhang-Wei Hong, Akhilan Boopathy, Pulkit Agrawal, Ila Fiete

Organizations: Massachusetts Institute of Technology

Abstract

Large-scale pretrained models are widely leveraged as foundations for learning new specialized tasks via fine-tuning, with the goal of maintaining the general performance of the model while allowing it to gain new skills. A valuable goal for all such models is robustness: the ability to perform well on out-of-distribution (OOD) tasks. We assess whether fine-tuning preserves the overall robustness of the pretrained model in image classification, and observed that models pretrained on large datasets exhibited strong catastrophic forgetting and loss of OOD generalization. To systematically assess robustness preservation in fine-tuned models, we propose the Robustness Inheritance Benchmark (ImageNet-RIB). The benchmark, which can be applied to any pretrained model, consists of a set of related but distinct OOD (downstream) tasks and involves fine-tuning on one of the OOD tasks in the set then testing on the rest. We find that though continual learning methods help, fine-tuning reduces robustness across pretrained models. Surprisingly, models pretrained on the largest and most diverse datasets (e.g., LAION-2B) exhibit both larger robustness losses and lower absolute robustness after fine-tuning on small datasets, relative to models pretrained on smaller datasets. We observe this collapse in contrastively pretrained (CLIP) models and their fine-tuned variants, where it grows with pretraining scale; the supervised models we test do not exhibit it. These findings suggest that starting with the strongest foundation model is not necessarily the best approach for performance on specialist tasks. https://jd730.github.io/projects/ImageNet-RIB

Figures & tables

Appendix figures & tables57 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Early Data Exposure Improves Robustness to Subsequent Fine-Tuning

    May 12, 2026Lawrence Feng, Gaurav R. Ghosal, Jacob Mitchell Springer +2Model Fine-TuningRetention

  2. Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models

    May 15, 2026Fabian Morelli, Arnas Uselis, Ankit Sonthalia +1Contrastive Language-Image Pre-Training ModelModel Fine-Tuning

  3. Robustness of Vision Foundation Models to Common Perturbations

    Apr 16, 2026Hongbin Liu, Zhengyuan Jiang, Cheng Hong +1Recent Vision Foundation ModelsVision Foundation Models