cs.LGJul 22, 2026

FedGuide: Diffusion Prior Alignment and Value Baseline Guidance for Heterogeneous Federated Reinforcement Learning

Authors: Zhilin He, Gauri Joshi

Organizations: Carnegie Mellon University

Abstract

Federated Reinforcement Learning (FRL) enables collaborative policy learning across distributed agents with heterogeneous environments. While recent methods based on variance reduction, divergence penalization, and momentum optimization improve FRL under heterogeneous settings, they still primarily synchronize policy or value-network parameters and do not explicitly address distributional mismatch among heterogeneous clients. Therefore, we propose \textbf{FedGuide}, a FRL framework that uses diffusion priors as behavior models to provide personalized data supported distributions for heterogeneous local policy learning. Instead of directly averaging local policies, FedGuide aggregates those diffusion priors through Optimal-Transport Mixture-of-Experts (OT-MoE), preserving heterogeneous behavior modes in distribution space. It further develops a Distribution Correction Estimation (DICE) value baseline to provide low-variance, return-aware guidance for local policy improvement. Experiments across heterogeneous environments show that FedGuide outperforms representative FRL methods in client-average returns, final-round performance, and worst-round robustness, while maintaining stable learning under stronger heterogeneity.

Figures & tables

Appendix figures & tables15 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Exploration-Driven Personalized Federated Reinforcement Learning via Intrinsic Motivation

    Aug 11, 2026Md Rafid Islam, Rafsan Jany, Zahid Hasan +1Federated LearningPrivacy

  2. FedQHD: Closed-Form Function-Space Federated Reinforcement Learning

    May 27, 2026Yuchen Hou, Yongshan Chen, Zhuowen Zou +4Federated Averaging