cs.LGMar 22, 2026

Federated Mixture-of-Experts Alignment on Mobile Edge Networks under Data Heterogeneity

Authors: Zihan Fang, Qianru Wang, Haonan An, Zheng Lin, Yiqin Deng, Symeon Chatzinotas, Yuguang Fang

Organizations: Hong Kong JC STEM Lab of Smart City and Department of Computer Science, City University of Hong Kong, Kowloon, Hong Kong SAR, China · School of Computer Science and Technology, Xidian University, Xi’an, China · Department of Electrical and Computer Engineering, The University of Hong Kong, Pok Fu Lam, Hong Kong, China · School of Data Science, Lingnan University, Tuen Mun, Hong Kong, China

Abstract

The growing demand for on-device large language model (LLM) services on mobile edge devices has driven the adoption of Mixture-of-Experts (MoE) architectures, which scale model capacity with limited computation. Since fine-tuning MoE-based LLMs relies on privacy-sensitive local data, federated learning (FL) offers a natural paradigm for collaborative training without exposing raw data. However, integrating MoE-based LLM fine-tuning into FL faces two critical challenges caused by data heterogeneity across clients: (i) divergent local data distributions drive clients to develop distinct gating preferences, so direct parameter aggregation yields a one-size-fits-none global gating network; and (ii) same-indexed experts develop disparate semantic roles across devices, leading to expert semantic blurring and degraded specialization. To address these challenges, we propose FedAlign-MoE, a federated aggregation alignment framework for edge computing systems that jointly enforces routing consistency and expert semantic alignment. Specifically, FedAlign-MoE aggregates gating behaviors by aligning routing distributions through consistency weighting and optimizes local gating networks through distribution regularization, maintaining cross-client stability while preserving discriminative local gating preferences. Meanwhile, FedAlign-MoE quantifies the semantic consistency of same-indexed experts across devices and selectively aggregates semantically aligned experts, ensuring stable and specialized global experts. Extensive experiments demonstrate that FedAlign-MoE outperforms state-of-the-art benchmarks, achieving faster convergence and higher accuracy in non-IID federated environments with lightweight computation and efficient communication.

Figures & tables

Explore similar work

CardsList
  1. Conflict-Aware Federated Fine-Tuning of Large Language Models with Mixture-of-Experts

    Jun 14, 2026Yijun Lu, Zihan Fang, Pengpeng Qiao +6Federated LearningMixture-Of-Experts

  2. AlignFed: Alignment-Aware Asynchronous Federated Fine-Tuning for Large Language Models in Heterogeneous Edge Environments

    Jun 6, 2026Yan Wang, Ziyi Gao, Rui WangEdge DevicesValue Alignment

  3. FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

    Jul 29, 2026Donghang Duan, Xu Zheng, Lizong Zhang +2HeterogeneityWeave