cs.LGSep 30, 2026

From Task Mixtures to Specialized Experts

Authors: Hojat Allah Salehi, Mehrdad Mahdavi, Andrew Arash Mahyari, M. Hadi Amini

Organizations: Florida International University, Miami, FL, USA · Security, Optimization, and Learning for InterDependent networks laboratory (solid lab), Miami, FL, USA · The Pennsylvania State University, University Park, PA, USA · Florida Institute for Human and Machine Cognition (IHMC), Ocala, FL, USA

Abstract

In collaborative foundation model fine-tuning, client data is rarely homogeneous. Instead, clients typically possess unknown mixtures of distinct data distributions, or tasks. Conventional federated learning primarily addresses heterogeneity across clients without explicitly resolving latent task mixtures within each client. We study this setting as compound heterogeneity, where data is heterogeneous both across and within clients. We study adaptation over a common frozen representation and show that, when tasks share the same feature geometry, the optimal model for a client's task mixture under squared loss is a convex combination of the optimal models for its underlying tasks. Thus, a single locally trained model represents the client's overall task mixture, while individual inputs may be drawn from different underlying task distributions. This motivates routing inputs to specialized experts, and we show that, when the task optima form a simplex, task-aligned routing achieves lower risk than any single adapted model for genuinely mixed clients. With access to a small set of task-labeled public samples, we derive a convex program to recover task experts and match them to their corresponding tasks. Our routing analysis shows that effective specialization requires input-dependent expert selection aligned with each client's task mixture. Motivated by this analysis, we propose FedSEE. Across our experiments, FedSEE avoids the negative transfer observed in the evaluated baselines and improves performance by 2.9 points overall and 3.7 points for the worst-served quartile.

Figures & tables

Appendix figures & tables15 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Federated Foundation Models Fine-Tuning with Heterogeneous Compressed Clients

    Jul 31, 2026Shengkun Zhu, Jinshan Zeng, Zhihua Allen-Zhao +5FedavgParameter-Efficient Fine-Tuning Methods

  2. FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

    Jul 29, 2026Donghang Duan, Xu Zheng, Lizong Zhang +2HeterogeneityWeave

  3. UB-SMoE: Universally Balanced Sparse Mixture-of-Experts for Resource-adaptive Federated Fine-tuning of Foundation Models

    May 15, 2026Van-Tuan Tran, Hong-Hanh Nguyen-Le, Marco Ruffini +1Heterogeneous Federated LearningMixture-Of-Experts Architectures