cs.LGOct 31, 2024

Local Superior Soups: A Catalyst for Model Merging in Cross-Silo Federated Learning

Authors: Minghui Chen, Meirui Jiang, Xin Zhang, Qi Dou, Zehua Wang, Xiaoxiao Li

Organizations: University of British Columbia · Vector Institute · Chinese University of Hong Kong · Meta

Abstract

Federated learning (FL) is a learning paradigm that enables collaborative training of models using decentralized data. Recently, the utilization of pre-trained weight initialization in FL has been demonstrated to effectively improve model performance. However, the evolving complexity of current pre-trained models, characterized by a substantial increase in parameters, markedly intensifies the challenges associated with communication rounds required for their adaptation to FL. To address these communication cost issues and increase the performance of pre-trained model adaptation in FL, we propose an innovative model interpolation-based local training technique called ``Local Superior Soups.'' Our method enhances local training across different clients, encouraging the exploration of a connected low-loss basin within a few communication rounds through regularized model interpolation. This approach acts as a catalyst for the seamless adaptation of pre-trained models in in FL. We demonstrated its effectiveness and efficiency across diverse widely-used FL datasets. Our code is available at https://github.com/ubc-tea/Local-Superior-Soups.

Figures & tables

Appendix figures & tables12 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. FedPLT: Scalable, Resource-Efficient, and Heterogeneity-Aware Federated Learning via Partial Layer Training

    May 4, 2026Ahmad Dabaja, Rachid El-AzouziHeterogeneous Federated LearningFederated Learning

  2. Federated Foundation Models Fine-Tuning with Heterogeneous Compressed Clients

    Jul 31, 2026Shengkun Zhu, Jinshan Zeng, Zhihua Allen-Zhao +5FedavgParameter-Efficient Fine-Tuning Methods

  3. Decoupled Training with Local Reinforcement Fine-Tuning in Federated Learning

    May 27, 2026Yuting Ma, Lechao Cheng, Xiaohua XuFederated LearningVision-Language Model Adaptation